Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
z7
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
z7
13d ago
François Chollet wrote in February that he expected ARC-3 to be saturated in "about one year". "Frontier models today perform very poorly with a minimal harness. However if big labs start directly targeting the benchmark like
2.
▲
by
z7
13d ago
Chollet writes he expects AGI now sooner than 2030, "given progress is happening faster than I expected." https://x.com/fchollet/status/2095607046129463577
3.
▲
by
z7
2mo ago
> The cost of generating the proofs for all 10 of these breakthroughs combined was under $2,000 at Sol API prices. https://x.com/polynoamial/status/2083470822258467194
4.
▲
by
z7
2mo ago
> The incredible Yitan Zhang ( https://newyorker.com/magazine/2015/02/02/pursuit-beauty ) worked on proving this conjecture for 7 years. Moh, his advisor, wrote that Zhang "failed miserably" i
5.
▲
by
z7
2mo ago
"Hey Leibniz, how do you live with yourself knowing that your binary system helped eventually replace human conversations?"
6.
▲
by
z7
3mo ago
"I just tested my hand in a mini version of this scanner. Images that are higher quality than MRI, whole body captured in <1 minute, virtually free to run. This is going to change medicine." https://x.com/Sebast
7.
▲
by
z7
7mo ago
"As Alexander predicted in 'AI 2027,' OpenAI did release a major new model in 2025; unlike in his forecast, it’s been a damp squib. Advances seem to be plateauing; the conversation in tech circles is now less about superintel
8.
▲
by
z7
8mo ago
The comparison isn't really like-for-like. NHTSA SGO AV reports can include very minor, low-speed contact events that would often never show up as police-reported crashes for human drivers, meaning the Tesla crash count may be drawing
9.
▲
by
z7
8mo ago
> The West is not complicit in the actions of the Iranian regime What about the 1953 CIA/MI6 coup that overthrew Iran's elected prime minister?
10.
▲
by
z7
10mo ago
"You only live once." Why state this as absolute fact? Seems a bit lacking in epistemic humility.
11.
▲
by
z7
11mo ago
Here's the Grokipedia submission (currently censored / flagged): https://news.ycombinator.com/item?id=45726459
12.
▲
by
z7
11mo ago
Hypothetically, what if the AI-generated blog post were better than what the human author of the blog would have written?
13.
▲
by
z7
1y ago
List of dates predicted for apocalyptic events: https://en.wikipedia.org/wiki/List_of_dates_predicted_for_ap...
14.
▲
by
z7
1y ago
Current cope collection: - It's not a fair match, these models have more compute and memory than humans - Contestants weren't really elite, they're just college level programmers, not the world's best - This doesn't
15.
▲
by
z7
1y ago
An encyclopaedia is a lossy representation of reality.
16.
▲
by
z7
1y ago
Meanwhile this new paper claims that GPT-5 surpasses medical professionals in medical reasoning: "On MedXpertQA MM, GPT-5 improves reasoning and understanding scores by +29.62% and +36.18% over GPT-4o, respectively, and surpasses pre-l
17.
▲
by
z7
1y ago
Yes, but the jump in performance from o3 is well beyond marginal while also fitting an exponential trend, which undermines the parent's claim on two counts.
18.
▲
by
z7
1y ago
>The actual benchmark improvements are marginal at best GPT-5 demonstrates exponential growth in task completion times: https://metr.org/blog/2025-03-19-measuring-ai-ability-to-com...
19.
▲
by
z7
1y ago
GPT-5 is #1 on WebDev Arena with +75 pts over Gemini 2.5 Pro and +100 pts over Claude Opus 4: https://lmarena.ai/leaderboard
20.
▲
by
z7
1y ago
Some previous predictions: In 2021 Paul Christiano wrote he would update from 30% to "50% chance of hard takeoff" if we saw an IMO gold by 2025. He thought there was an 8% chance of this happening. Eliezer Yudkowsky said "at
21.
▲
by
z7
1y ago
How do you explain Grok 4 achieving new SOTA on ARC-AGI-2, nearly doubling the previous commercial SOTA? https://x.com/arcprize/status/1943168950763950555
22.
▲
by
z7
1y ago
"Grok 4 (Thinking) achieves new SOTA on ARC-AGI-2 with 15.9%." "This nearly doubles the previous commercial SOTA and tops the current Kaggle competition SOTA." https://x.com/arcprize/status/1943
23.
▲
by
z7
1y ago
Quoting Chollet: >I have repeatedly said that "can LLM reason?" was the wrong question to ask. Instead the right question is, "can they adapt to novelty?". https://x.com/fchollet/status/18663
24.
▲
by
z7
1y ago
It's just predicting tokens: https://old.reddit.com/r/singularity/comments/1jl5qfs/its_ju...
25.
▲
by
z7
1y ago
Why are you hallucinating feelings? Also, appeal to authority. ("Why are your feelings relevant to the wizarding laws of Hogwarts?")
26.
▲
by
z7
1y ago
>For starters, this completely blocks generation of anything remotely related to copy-protected IPs It did Dragon Ball Z here: https://old.reddit.com/r/ChatGPT/comments/1jjtcn9/the_new_im... Rick and
27.
▲
by
z7
2y ago
The beginning of a new kind of discrimination - call it 'synthetic racism.' AI-generated music is being dismissed outright even before listening to it, not based on quality or enjoyment but purely on its artificial origin. Just as
28.
▲
by
z7
2y ago
Waymo's driverless taxis are currently operating in San Francisco, Los Angeles and Phoenix.
29.
▲
Competitive Programming with Large Reasoning Models
(arxiv.org)
6 points
by
z7
2y ago
|
0 comments
30.
▲
by
z7
2y ago
I don't understand it either. Why is Elon Musk a "terrorist"? And why is this the most upvoted post? Maybe being European limits my ability to comprehend American political rhetoric.
More ›