Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
eckr
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
eckr
16d ago
"Claude Mythos 5.1 is identical to Fable 5.1, but it offers more permissive safeguards for vetted individuals and organizations" Then why does it have separate datapoints for Terminal Bench, and score higher? Something doesn'
2.
▲
by
eckr
1mo ago
Maybe this is just my experience, but have people had trouble with 3.6 Flash just... getting things it has seen in its context correct? I don't know if it's been insanely benchmaxxed or what, but it'll pull information from w
3.
▲
by
eckr
2mo ago
There was a startup that did this for Llama 3, I forgot their name. Etched is also doing some similar things I believe.
4.
▲
by
eckr
2mo ago
Obviously it comes down to that, but you can't make the claim that GPUs aren't a huge part of it. Otherwise, billions wouldn't be getting invested into them in the west, no? And "financial backing" essentially boils
5.
▲
by
eckr
2mo ago
This is fair, with the caveat that we don't know for how long this model has been around internally in China either. So we can only go about appearance / releases.
6.
▲
by
eckr
2mo ago
This, essentially. Especially in compute I don't think it's debatable they have less resources than the frontier labs in the west.
7.
▲
by
eckr
2mo ago
These benchmark numbers are insane. The days when China was 6 months behind are over? How are they doing this with so much less resources than the US??? I have so much respect for the researchers there
8.
▲
by
eckr
2mo ago
Because, realistically that's all programmers ever need, would be my guess. I do think linear algebra is an extremely interesting topic in its own right / outside, but yeah.
9.
▲
by
eckr
3mo ago
In the past, they just ran Deepseek OCR on your image and extracted the text, then gave it to a language only model. I believe now there is a model that actually takes images as input directly.
10.
▲
by
eckr
5mo ago
I don't think this is private knowledge guessing from when and how I was told, so I feel comfortable sharing it. When I talked to some Huawei representatives, I was told DeepSeek V4 was trained entirely on Huawei chips. It's up t