Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gbnwl
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
24 ms
·
1.
▲
by
gbnwl
10d ago
This is a nit but his name is actually Erik Satie not Eric Satre.
2.
▲
by
gbnwl
28d ago
It seems like the entire thought process you’re trying to sell hinges on the idea that OpenAI reported the attack first. Did you forget that it was actually HuggingFace that reported it first, and OpenAI only stepped forward latter?
3.
▲
by
gbnwl
1mo ago
It’s just a boring and unhelpful complaint that afaict largely serves to soothe the commenters ego rather than point at anything insightful that’s useful or predictive. Point me to your favorite “next token predictor” comment that was actua
4.
▲
by
gbnwl
1mo ago
I just prefer HN comments to be better reflections of reality. There is an unspoken expectation here that people here know what they’re talking about especially when it comes to technical matters. The rise of LLMs has given way to a HN bran
5.
▲
by
gbnwl
1mo ago
Fair so let me be clear. I’m whining because the “next token predictor” reductionist point of view has been wrong and is only growing more wrong with time. Clearly these things can do things that actually matter. Do you disagree?
6.
▲
by
gbnwl
1mo ago
Are they useful or not? Will they continue changing the world or not? People who choose one way or the other for describing them typically fall on one side or the other in these questions imo. What do you think? Will these next token predic
7.
▲
by
gbnwl
1mo ago
Every day I wake up and open HN. “LLM has made legitimate mathematical discoveries” —> Wow the rate of progress is amazing. Highly upvoted. “LLM does something not good” -> Does everyone else not realize LLMs are just dumb next token
8.
▲
by
gbnwl
2mo ago
There are articles with far fewer upvotes and comments ranking higher on the front page right now, despite being the same age or older than this one. HNs opaque ranking system at it again.
9.
▲
by
gbnwl
2mo ago
Everyday? Which 10 problems were solved by mathematics grad students in the past 10 days? OK I’ll grant that it’s not your obligation to be my search function (despite you making the wild assertion in the first place), so instead can you ju
10.
▲
by
gbnwl
2mo ago
I think the fundamental difference between our assumptions is you believe prompts to be optimized for tasks rather than model-task pairs. The only elaboration I can give you is empirical observations and model providers own guidance (as som
11.
▲
by
gbnwl
2mo ago
Experienced similar between 5.4-mini vs 5.6-luna in our own pipelines but after spending some time on prompt optimization and testing out various reasoning effort levels 5.6-luna was well worth it. Did you just replace model selection while
12.
▲
by
gbnwl
2mo ago
We can assume (outside of Ollama) that they meant the strongest model from each lab. If you limit yourself to just looking at the literal strings in the list, literally none of these are models. What model is "Deepseek" or "G
13.
▲
by
gbnwl
2mo ago
How is this any different than what we have already? We've had this ability for ages (6+ months, decades in the AI world), you can literally today easily prompt CC or Codex to use subagents to accomplish tasks and they'll do it w
14.
▲
by
gbnwl
3mo ago
As usual HN posters are hyper aware of other's credentials while ignoring that their BS in CS (if that) doesn't magically qualify them to assess everything in every domain. "I'm a software engineer, I'm sure if I ha
15.
▲
by
gbnwl
3mo ago
https://xcancel.com/trq212/status/2014051501786931427#m
16.
▲
by
gbnwl
3mo ago
Hm first Shazeer and now Jumper, DeepMind getting hollowed out this week.
17.
▲
by
gbnwl
4mo ago
Don't forget that the denominator (total number of outstanding shares) will be increased by this as well. So even if the market cap reacted exactly one to one like you're proposing the per share price wouldn't stay constant n
18.
▲
by
gbnwl
4mo ago
No they didn't, they predict they'll get that much. Also worth noting the prediction assumes running at MXFP4/FP8 quantization.
19.
▲
by
gbnwl
4mo ago
Frontier as in "Frontier Model" is a legitimate vocabulary term you should probably be aware of in 2026. It's not something the author made up or chose randomly, it's common parlance in the space.
20.
▲
by
gbnwl
4mo ago
Can you really look yourself in the mirror and say with a straight face that fundamentally nothing has changed about the relationship between the US and its allies? Do you really think Europeans will be quick to forgive the wrongs of this a
21.
▲
by
gbnwl
4mo ago
Probably a better way to phrase it would be keeping “pace“. Yes, they are still behind but by about the same amount as they always have, they aren’t drifting further behind. Like two marathon runners, one a a mile ahead, but both maintainin
22.
▲
by
gbnwl
4mo ago
I mean yes? Considering that we're here reading the news that they've agreed to this.
23.
▲
by
gbnwl
4mo ago
What exactly is the intuition behind taking something inherently linear like a sequence of 100 days and presenting it as a graph with no information given about the rationale or reasoning behind the edges.
24.
▲
by
gbnwl
5mo ago
The symbol is not the thing. The map is not the territory. Ceci n'est pas une pipe.
25.
▲
by
gbnwl
5mo ago
Of course the motivation makes sense on the surface. What I'm getting at is that the supply of capital vs the supply of potential "control of the future" plays feels incredibly imbalanced. Money seems to be so desperate to mo
26.
▲
by
gbnwl
5mo ago
Not the first to notice this I'm sure but it feels like there's an insane amount of pressure pushing capital towards anything with a hint of AI legitimacy. It's as if asset owners across the planet have come to a consensus th
27.
▲
by
gbnwl
5mo ago
The entire post reads like it was generated via LLM as well.
28.
▲
by
gbnwl
5mo ago
Never thought I'd see the day ragebait made it to HN. Yes, let's pretend doing a long jump on the moon is comparable to running a marathon at its prescheduled time at its prescheduled location. Weather is always a factor in sports
29.
▲
by
gbnwl
5mo ago
I didn't express this well but my interest isn't "who is in the top spot", and is more _why and _how various labs get the results they do. This is also magnified by the fact that I'm not only interested in hosted pr
30.
▲
by
gbnwl
5mo ago
I’m deeply interested and invested in the field but I could really use a support group for people burnt out from trying to keep up with everything. I feel like we’ve already long since passed the point where we need AI to help us keep up wi
More ›