Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sognetic
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
1.
▲
by
sognetic
12d ago
There are a bunch of approaches that do this kind of thing to reduce token usage ("semble" came to mind, technically different but functionally similar) but their performance is usually mixed because the models haven't been R
2.
▲
by
sognetic
3mo ago
Enterprises still have big contracts with github, those companies are imposing tight spending limits now and if the open weight models enable those limits to last a bit longer that's probably quite popular.
3.
▲
by
sognetic
8mo ago
Everything is currently pointing towards inference being the main cost driver for LLMs in the future. Test-time-compute requires huge amounts of tokens in inference and makes providing frontier models as services unprofitable. Anyone not un
4.
▲
by
sognetic
8mo ago
Interesting! So did you do any experiments on a relevant subset of the data to test whether LLM performance degrades by introducing a new, presumably unknown to the LLM, format?
5.
▲
by
sognetic
1y ago
Accepting the possibility of committing the old "solving social problems with technological solutions" fallacy: I wonder if an offtopic channel without history (or only a very limited one) could help here. Something that prevents