Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
iagooar
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
iagooar
16d ago
As much as I love my CLI tools, I just cannot go back to not having native built-in browser integration, links, etc which the (Mac) desktop apps like ChatGPT or Claude offer. I do use Hermes sometimes, but it tends to keep growing skills an
2.
▲
by
iagooar
19d ago
I am also seeing slower speeds, roughly the same ballpark, sometimes even lower - 10-11 tok/s on M5 Max. If there is that ONE version (GGUF or MLX) that runs roughly as fast as 3.6 used to run, please let me know. What would be incredi
3.
▲
by
iagooar
27d ago
My observation is that people who are smart (or think they are) tend to express themselves via trying to be overly correct - on everything. And that gets exhausting. People who try to be too precise and accurate, pay it with their happiness
4.
▲
by
iagooar
1mo ago
10 years. I still have the game in my Steam library, waiting for me to try it. I guess I will never try it, once you hit your 40s, games lose their appeal anyway.
5.
▲
by
iagooar
1mo ago
I am working on a framework that allows creating very complex SaaS products with AI, that actually work. Very complex = +100k LOC of one-shotted AI-generated code. Codename o2p - outcome-to-proof. It is a research-first method for designing
6.
▲
by
iagooar
1mo ago
I love DeepSeek V4 Flash since the pre-0731, now even more. It is the first model that is truly too cheap to meter. But I find it having a pretty significant problem with tool calling - no idea why, but tool calling with it is SLOW. As long
7.
▲
by
iagooar
1mo ago
Not sure I agree. It might sound like theater if you expect it to solve your security issues. But it is extremely good at finding undocumented or broken processes - these processes tend to hide the biggest offenders.
8.
▲
by
iagooar
1mo ago
So far every single test I've done has worked correctly, I put a lot of work and time into preparation and planning, each workflow was designed, researched and optimized before implementation. The specification was already a massive do
9.
▲
by
iagooar
1mo ago
No, LLMs have not plateaued, and each of the latest releases has been a proof of that. Fable 5 is a model you instantly FEEL how smart and superior it is. GPT 5.6 Sol is a HUGE incremental improvement in multiple directions and dimensions.
10.
▲
Ask HN: I had Codex and GPT 5.6 Sol running for 12 days, 870k+ LOC. Now what?
1 points
by
iagooar
1mo ago
|
3 comments
11.
▲
by
iagooar
2mo ago
Having invested in a machine with 128GB of RAM, I would love seeing something a bit larger than 27B / 35B, possibly a 54B dense model or 70B MoE would be much closer to the Qwen 3.8 Max experience.
12.
▲
by
iagooar
2mo ago
What is your story, how did you get into robotics? Was it your first choice?
13.
▲
by
iagooar
2mo ago
I need to use this thread to mention the Polish payment system Blik. It is one of the most convenient payment systems out there - you generate an ephemeral 6-digit code in your banking app, copy it to the vendor, confirm, done. Works P2P wi
14.
▲
by
iagooar
3mo ago
Scanning mailbox, reading and classifying emails. Scanning a knowledge base, reading and improving individual articles, reading support interactions and creating summaries, checking and researching new leads that signed up. So much that is
15.
▲
by
iagooar
3mo ago
On paper the M4 should be roughly 1/3 of the M5, in practice it is only 1/2. With the right, optimized model like qwen3.6 35B MoE MLX you can get over 40 tok / sec on it. I run dozens of background jobs that are not time-crit
16.
▲
by
iagooar
3mo ago
I am not going to flag you, I am much OK with having good arguments. I just purchased a Mac Mini M4 Pro 64GB for $3k - 2nd hand of course. I am not a hater of Nvidia and I am planning on building a workstation based on RTX cards. You clearl
17.
▲
by
iagooar
3mo ago
Buy a refurished or 2nd hand one.
18.
▲
by
iagooar
3mo ago
I disagree LAN connection is the bottleneck. I do even work with it remotely via Tailscale on shaky hotel WIFI and it works fine (or as fine as any other API-based model).
19.
▲
by
iagooar
3mo ago
Just buy a Mac Mini really is good advice if you want to get into real, always-on convenient agentic work. Soon it is going to be good even for coding using local LLMs. Until then, just run API models on it for coding, local LLMs for "
20.
▲
by
iagooar
3mo ago
I thought they might ship an M5 Max version, but you are probably right.
21.
▲
by
iagooar
3mo ago
qwen3.6 27B MLX 8bit -> 15 tok / sec. A bit slow but it is a delightful model to use, and smart too. qwen3.6 35B A3B MLX 8bit -> 85-90 tok / sec! It is impressively fast and roughly 90% as good as 27B (in my opinion).
22.
▲
by
iagooar
3mo ago
My problem is I won't accept anything lower than the 96GB the RTX Pro 6000 Blackwell has. My dream is a workstation with 2x Pro 6000 to run DeepSeek v4 Flash comfortably, possibly qwen 3.6 / ornith on turbo speed. But man, I have
23.
▲
by
iagooar
3mo ago
Ballpark 25-30 tok / sec on the Mac Mini Pro M4 + qwen3.6 35B. The generation itself is good, prefill is known to be slow on any Apple M-chip architecture. It is really decent.
24.
▲
by
iagooar
3mo ago
M5 Max. But I also have a MacMini M4 Pro 64GB. Qwen3.6 runs on the M4 just fine - sure the M5 is at least 2x the speed. If Apple launches a MacMini with an M5, I will be the 1st one to get it.
25.
▲
by
iagooar
3mo ago
Get a 2nd hand one. I was lucky enough to get a new one first, last week I get a 2nd hand one in order to run one of my Hermes minions at work.
26.
▲
by
iagooar
3mo ago
I love my MacBook Pro M5 128GB RAM and I love qwen3.6. BUT DO NOT buy this MacBook if you plan on doing serious coding using local LLMs with it. The reason is simple: your fingers will burn and your head will explode from the noise. Running
27.
▲
by
iagooar
3mo ago
Then it probably was wishful thinking on my side...
28.
▲
by
iagooar
3mo ago
I guess people are tired of each instance of an Electron-based app using 1GB+ of RAM.
29.
▲
by
iagooar
3mo ago
I have noticed that Opus and GPT 5.5 are very good at adjusting their thinking / reasoning intensity depending on the task at hand, something the open weights models are still not as good at. In addition to that, some of the open weigh
30.
▲
by
iagooar
3mo ago
It is a combination of Hermes agent as the orchestrator and a custom extractor script (that uses qwen or any other LLM) that runs every 2 mins (on Mac via launchd). I had the code + skill written by Hermes. The beauty of it is that Hermes i
More ›