Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
bermudi
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
bermudi
14d ago
Muse 1.2 wrote a terrible "smart summaries" extension for my pi setup. It was sending every single steamed chunk for summarization instead of waiting for the full CMD. This is an error I would expect from sonnet 4, not a model tha
2.
▲
by
bermudi
14d ago
I honestly can't believe serious people are making this argument on a straight face. Gemini 3.7 flash outputs so many tokens per answer it doesn't matter how fast its TPS is, sol will end up being both cheaper and faster than Gemi
3.
▲
by
bermudi
1mo ago
This only makes me understand how flawed AAII is. This Qwen model is nowhere close to the other models in that score range.
4.
▲
by
bermudi
2mo ago
If you're working on smol? How complex is your work? The agent has to do everything using sed? Did you write your own tools? I guess my question is, why aren't you using pi? The difference in tokens between the two also makes supe
5.
▲
by
bermudi
2mo ago
DeepSeek being DeepSeek. v3 and R1 went over the same and had multiple versions
6.
▲
by
bermudi
3mo ago
Source? The most trusted benchmark right now (deepSWE) scores better or just as well on their minimal harness than when using CC or codex
7.
▲
by
bermudi
3mo ago
I wonder if you're as cynical and untrustworthy of American companies as well or is it more of a racism kinda thing
8.
▲
by
bermudi
4mo ago
Everything is more expensive than deepseek. They aren't frontier in intelligence but they are the frontier in cost per intelligence
9.
▲
by
bermudi
3y ago
What surprised me the most was the fact that people have enough disposable income to pay for search to make this viable