Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nxtfari
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
nxtfari
8d ago
We’re reaching quadratic slop. Slop projects that don’t understand what they’re shipping built on top of slop projects that also don’t understand what they’re shipping. Magnificent.
2.
▲
by
nxtfari
13d ago
Good feature but god I hate there’s a whole new keyword for a minor modification on an existing concept. Very unnecessary.
3.
▲
by
nxtfari
20d ago
As the saying goes every 5 minutes in Silicon Valley another company reinvents ROS
4.
▲
by
nxtfari
22d ago
Agree, I remember when even half precision made its way into C# sometime around 2020 (I didn’t know much about ML then) and I thought, well I guess that’s a worthwhile tradeoff but I can’t imagine going lower. Lo and behold (1-bit Bonsai) h
5.
▲
by
nxtfari
1mo ago
Sure, then you should allow that they are also tools for transforming sequences of questions into answers. Language models are based on compression of information, using them as a knowledge base is entirely within capability.
6.
▲
by
nxtfari
2mo ago
Surprisingly incoherent for Anthropic and Dario (cue peanut gallery — “always has been!” No, I don’t think so. I think this is new). It seems to me like there is just no good answer to how one could possibly stop open weight models from bei
7.
▲
by
nxtfari
2mo ago
They must have done the math to show even if you peg Kimi K3 generating 24 hours a day for an entire month it doesn’t exceed 10k in opex? Very interesting.
8.
▲
by
nxtfari
2mo ago
Jagged frontier is not the same as being benchmaxxed. Benchmaxxed is à la Goodhart's Law "when a measure becomes a target, it ceases to be a good measure." Jagged frontier is about how models that seem superhumanly intelligen
9.
▲
by
nxtfari
2mo ago
Haha, this is awesome. Hadn't heard of it, thank you for the link.
10.
▲
by
nxtfari
2mo ago
If you haven’t been really running and testing these models yourself, they are all benchmaxxed. No matter how close they score to frontier on whatever metric, they always fall apart in real world tasks and their token efficiency is ridiculo
11.
▲
by
nxtfari
2mo ago
Very high quality and info-dense article. Cheers.
12.
▲
by
nxtfari
2mo ago
This is really smart, like the author said, old idea but cleverly applied. In case anyone wants a summary: don’t one shot, don’t use plan mode and hand off the plan to cheap executors, ask the frontier model to explore, create a todo list,
13.
▲
by
nxtfari
2mo ago
Everyone using Claude Fable to verify this proof is so funny. If you read the definition of the Jacobian Conjecture and (I am not exaggerating this) have passed a college Calc 3 class, you can just verify the proof yourself in 30 seconds. T
14.
▲
by
nxtfari
2mo ago
Eh, Minimax M2.7 also took a similar amount of time (actually longer) between availability and weights release.
15.
▲
by
nxtfari
2mo ago
you could live (100k a year) with the interest of 0.004% of that in the bank.
16.
▲
by
nxtfari
2mo ago
Apple (and the ARM ecosystem as a whole) has never really needed massive GPU compute before, it’s always been about power efficiency and just enough GPU oomph to make UI fluid. Even historic Mac Pro workloads never really needed tons of GPU
17.
▲
by
nxtfari
3mo ago
I think we should make it illegal to not specify the quantization in the headline for these types of posts.
18.
▲
by
nxtfari
3mo ago
this is really cool but it seems very unlikely that someone targeting an exotic system not supported by rust (mostly embedded and ancient mainframe targets) would be willing to trust a beta transpiler to not inject any bugs or leaks in the
19.
▲
by
nxtfari
3mo ago
An important idea I’ve observed true across industries is “prices rise like a rocket and fall like a feather,” meaning that even though price rises are usually genuinely driven, you can bet that once they’re up the involved parties are doin
20.
▲
by
nxtfari
3mo ago
this is really clever, props
21.
▲
by
nxtfari
3mo ago
Location: NYC / LA Remote: yes Willing to relocate: NYC or LA Technologies: Motion planning, autonomy, embedded microcontrollers, embedded Linux, perception systems, controls. C, C++, Python, Rust Résumé/CV: https://git
22.
▲
by
nxtfari
3mo ago
they’re doing m7 on the intel 18a fab, which is exactly that
23.
▲
by
nxtfari
3mo ago
One of the stupidest things about this is we talk all day along about how frontier models don’t just interpolate distribution, then can extrapolate out. Then something like this comes along and a model can generate gore or CSAM so therefore
24.
▲
by
nxtfari
3mo ago
My honest read is that, having everything — the data centers, the compute, the models (however misaligned they might be), the only thing xAI is missing is users. They don’t have any users because the only people who use Grok are essentially
25.
▲
by
nxtfari
4mo ago
this being HN, from the title i genuinely had no idea whether this link would be about music, the apple graphics acceleration framework, or ore deposits.
26.
▲
by
nxtfari
4mo ago
really impressive. did not expect this from infineon.
27.
▲
by
nxtfari
4mo ago
> One of the lessons of philosophy is that once you adopt any particular value system, almost all philosophers either become immoral or caught up in meaningless and trivial quibbles. Can you explain more about this?
28.
▲
by
nxtfari
5mo ago
> Do people have extremely complex Actions that I can't fathom? Yes. Think CI jobs that test every candidate PR against a matrix of build targets, run fuzzing, run simulation tests, run bench regression tests, etc etc. Modern CI wor
29.
▲
by
nxtfari
5mo ago
I had issues with Qwen thinking endlessly when I didn’t know I wasn’t using the temp/top_k/min_p/etc settings specified in the readme. I’ve never had an issue with Gemma 4 thinking endlessly but could possibly be the same.
30.
▲
by
nxtfari
5mo ago
This makes a lot of my experience with Qwen make sense. I’ve watched all the benchmarks imply how close it should be to various GPT or Claude releases, but in my own use chatting with it or trying to get it do agentic tasks it was nowhere n
More ›