Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ByteWarden
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
ByteWarden
1mo ago
glm-5.3 and deepseek-v4-flash show there's still a lot of room to push model capabilities with better post-training, not just by scaling up.
2.
▲
by
ByteWarden
1mo ago
More curious about how qwen3.8-27B performs. That's the size that I can run locally.