Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
inferencecoder
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
inferencecoder
1mo ago
I meant serious players often do serve on a single node. They can beat API pricing as well. Multi-node can add gain, but also adds a lot of deployment complexity so its just not always possible or optimal.
2.
▲
by
inferencecoder
2mo ago
> Not to mention no one serious is serving this on 8xB200 instead of multiple nodes: the vast majority of Moonshot's inference work is focused on PD-disaggregation The GPU price discourse is absurd, but many are serving models on si
3.
▲
by
inferencecoder
2mo ago
To begin with, as I pointed out separately, the price comparison that makes the headline is absurd. No one is easily getting the MI355X's at $2.5/hr. And then the eval is unreplicable and poorly defined.
4.
▲
by
inferencecoder
2mo ago
I think it's fair to expect an extensive review of an article before publishing. Not everything has a set of serious flaws.
5.
▲
by
inferencecoder
2mo ago
https://chatgpt.com/share/6a6f09ff-2830-83ea-9578-de3016cfae...
6.
▲
by
inferencecoder
2mo ago
Wafer is making themselves synonymous with slop in the inference space. Exaggerated unfair comparisons in all their results, twitter hype posts with alarm emojis etc. > $2.50/GPU-hr for the MI355X, $6.00 for the B300, and $4.25 for
7.
▲
by
inferencecoder
2mo ago
There is barely any effort, a simple GPT5.6 sol pro query rips the post apart.