Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nearbuy
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
nearbuy
4d ago
Buckmaster (the mathematician) and Alpöge used and credit AI substantially for their proof. Even if OpenAI did copy their ideas, it still wouldn't show that this didn't come from AI improving. OpenAI's proof is substantially
2.
▲
by
nearbuy
7d ago
You're conflating two very different meanings of "the speed of light". Confusingly, "speed of light" can refer to the universal constant, c, which does not change in glass, or to the speed that light travels in a me
3.
▲
by
nearbuy
10d ago
Why? 1. I didn't say LLMs have made any breakthroughs in math, not because they haven't, but because it's irrelevant to my point. The parent comment is using the same argument academic research opponents have long used agains
4.
▲
by
nearbuy
10d ago
There are several hundred thousand mathematicians producing hundreds of thousands of new results in math each year. Why can't most people name any human contributions to mathematics from the past decade? What is the finished result tha
5.
▲
by
nearbuy
12d ago
The Kevin Buzzard post linked at the top says they budgeted £1M over 5 years for a smaller proof.
6.
▲
by
nearbuy
13d ago
I think you're misunderstanding. Astra is at the top of the official ARC-AGI leaderboard, with an ARC-AGI approved harness. It's not a harness specialized for ARC-AGI. It just does the same thing the regular ChatGPT interface does
7.
▲
by
nearbuy
13d ago
People assume that's the reason because it's intuitive and "strawberry" is one token. But that doesn't explain why those models would also often get it wrong for "StRaWbErRy" or even "s-t-r-a-w-b-e-r-
8.
▲
by
nearbuy
13d ago
I don't think we should count the lower tier models if we're discussing what the top ones are capable of. No one was suggesting that Sonnet is AGI.
9.
▲
by
nearbuy
13d ago
The last version to fail on those questions was GPT 4.5. Meanwhile most humans fail to correctly answer how many f's are in the sentence, "Finished files are the result of years of scientific study combined with the experience of
10.
▲
by
nearbuy
1mo ago
It also looks like they're saturating the test, with one LLM hitting the maximum possible score. ( https://www.trackingai.org/home ) The test wasn't made to accurately measure IQs that high.
11.
▲
by
nearbuy
1mo ago
The paper also fails to show that their central example, Einstein, relied on sensory experience for his intuition leaps rather than general reasoning. They just kind of claim that thought experiments require sensory experience. But you can
12.
▲
by
nearbuy
1mo ago
Only a tiny, tiny fraction of the parameters are encoding information that's specific to a particular programming language. Even if you could remove those without degrading performance, it would have a negligible effect on the model si
13.
▲
by
nearbuy
1mo ago
The problem is Claude Fable is now better than most programmers I know at software architecture and performance optimization as well.
14.
▲
by
nearbuy
2mo ago
That's a good reason to use a dishwasher, and you should keep doing it. But the overall waste is small and people aren't going to care. I'm not saying people who already have a dishwasher will throw it away. But a robot makes
15.
▲
by
nearbuy
2mo ago
Sure, but then there's no such thing as a network that isn't a classifier. Every physically computable function that terminates in finite time will map an input to a fixed set of outputs. And it goes against the common usage, wher
16.
▲
by
nearbuy
2mo ago
The LLM has processed two data modalities derived from the apple (text and vision). Your brain processed a third (taste). But it is still just a data stream, sensing compounds and chemical properties of the apple and turning it into a strea
17.
▲
by
nearbuy
2mo ago
LLMs are not classifiers. A classifier is an algorithm or neural net that assigns a label from a fixed set of labels to an input. You can broaden the definition of classifier to anything that internally divides its input space into regions,
18.
▲
by
nearbuy
2mo ago
I have a dishwasher and I still usually just hand wash. It takes about 10 seconds to wash a dish. The side benefit is all your dishes are always available. With the dishwasher, up to one full dishwasher load are dirty at any time, which mea
19.
▲
by
nearbuy
2mo ago
Unitree's R1 humanoid robot is only about $6000, and it's still a nascent, smallish scale technology. They will come down. If the future home robots are any good, it saves you from buying a dishwasher and robot vacuum. It can repl
20.
▲
by
nearbuy
2mo ago
Yes, much like that. If he hadn't destroyed evidence, he could have argued it was malicious prosecution.
21.
▲
by
nearbuy
2mo ago
What you're describing is malicious prosecution or abuse of process. It's illegal and it would destroy the prosecution's case. Not only that, but the victim could sue for damages.
22.
▲
by
nearbuy
2mo ago
The irony is in this case the in-context and classifier "guardrails" would have almost certainly stopped the attack while their attempts at your definition of guardrails (the sandboxing) failed. In general, people keep trying to m
23.
▲
by
nearbuy
2mo ago
It's a strange experiment. Claude and GPT aren't generating the video. They're directing and editing it, and they request video from a generative video model using mainly text-to-video. Neither Claude nor GPT can actually wat
24.
▲
by
nearbuy
2mo ago
The UNESCO/World Bank literacy rate is basically defined how you thought. But high income countries don't usually report this because literacy by this measure is nearly universal. So they often report at higher thresholds (e.g. ho
25.
▲
by
nearbuy
2mo ago
Questions like that cost a tiny fraction of a cent. "What's the capital of Sri Lanka?" cost a fifth of a cent at GPT 5.5 API price, and would cost a fraction of that if the question were routed to a more suitable, cheaper mod
26.
▲
by
nearbuy
3mo ago
The usage is irrelevant if we're interested in cost per token. If you use it half as much, you get half as many tokens at half the cost. It's still $5.56 in electricity per million output tokens either way (using $0.20/kWh, a
27.
▲
by
nearbuy
3mo ago
Not much point in serfdom when they don't have use for human labor.
28.
▲
by
nearbuy
3mo ago
> I feel like you're also doing something weird—sorta strawmanning and sorta being conveniently inconsistent. You're reframing his argument as something much softer. I'm going to firmly push back on this. A reasonable, str
29.
▲
by
nearbuy
3mo ago
You're not really engaging with PG's argument so much as nitpicking around the edges. Are people becoming billionaires primarily by cheating and stealing, or by making something people want? That's the core argument. Were Git
30.
▲
by
nearbuy
3mo ago
Anthropic's plan may not have great odds, but your proposal is orders of magnitude worse. There are plenty of groups opposing AI. They don't get billions in investment. They don't have a clear way to stop OpenAI or Qwen. They
More ›