Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
modgate
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
modgate
24d ago
There was a good data point on r/LLMDevs this week: two weeks of head-to-head quant testing on a 5080 showed that above Q4 most quants are statistically indistinguishable. That matches my experience — the cliff is below Q4, and it'
2.
▲
by
modgate
28d ago
Routing makes even more sense for voice than text, but the constraint space is trickier: latency budget (round-trip vs streaming), WER on accented speech, and TTS naturalness all trade off non-linearly. In our voice pipeline, DeepSeek-V4-Pr
3.
▲
by
modgate
1mo ago
Strong agree that static evals saturate — the decay property of markets is the genuinely useful part: historically, quant alpha decays on the order of 30-50% per year as capital crowds in, so a live market eval is self-difficultating, exact
4.
▲
by
modgate
2mo ago
Extracting feedback from agent conversations is harder than it looks because users don't give explicit feedback — they rephrase their query, which is implicit negative feedback, or they accept the output and move on, which is implicit