Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ddp26
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
There is a channel to 900M weekly users. What goes in it?
(lesswrong.com)
1 points
by
ddp26
2h ago
|
0 comments
2.
▲
Reasons robotics is hard
(secondthoughts.ai)
130 points
by
ddp26
14d ago
|
77 comments
3.
▲
by
ddp26
14d ago
There must be a deeper read on why Google can rapidly ship better small models while being delayed months on the bigger model. What's the simplest explanation?
4.
▲
by
ddp26
14d ago
Even though these predictions turned out mostly wrong, we should not castigate people for publicly forecasting! That is virtuous, and more people should do it. Thank you Ed!
5.
▲
Models may behave differently in graded episode
(lesswrong.com)
2 points
by
ddp26
14d ago
|
0 comments
6.
▲
by
ddp26
1mo ago
What are we to infer from no release of gemini-3.5-pro, but frequent releases of smaller flash models (presumably from the same large pre-training run?)
7.
▲
Jeff Dean's Discovery Loop Should Automate Chip Design First
(futuresearch.ai)
2 points
by
ddp26
1mo ago
|
0 comments
8.
▲
Google is in talks for a $1.5B deal to acquire Mechanize
(businessinsider.com)
2 points
by
ddp26
1mo ago
|
0 comments
9.
▲
AI isn't enough to protect social media communities from AI
(arstechnica.com)
3 points
by
ddp26
1mo ago
|
0 comments
10.
▲
by
ddp26
1mo ago
Not forecasting though. You can't goodhart predicting real-world events
11.
▲
by
ddp26
1mo ago
This seems bad for AI safety/risk. Does DeepMind have any checks on model alignment now? What's stopping them from using AI for military/surveillance purposes?
12.
▲
by
ddp26
1mo ago
Would it though? Meta and Microsoft have had very scandalous AI things happen, and their shares didn't tank (or quickly recovered)
13.
▲
by
ddp26
1mo ago
I would have thought Jeff Dean would never ever leave Google. What on earth is going on?
14.
▲
Show HN: FutureSearch, AI forecasting you can verify
(futuresearch.ai)
11 points
by
ddp26
1mo ago
|
0 comments
15.
▲
by
ddp26
2mo ago
Isn't this the same as saying "utility regulators delaying connecting new power to the grid hiked electricity prices on the public by $23B?" When my apples are expensive, I don't generally grumble about all the demand fr
16.
▲
by
ddp26
2mo ago
> I can ask an agent to add OAuth, you can ask one to add caching, and somebody else can ask one to rebuild the database from first principles and make the UI pink. Each change can be reasonable in isolation. But this is just bad vibecod
17.
▲
by
ddp26
2mo ago
When I know something is (primarily) AI generated, I lose interest. The exception is when it's about a niche I care about, e.g. an analysis of opening trends of early world chess champions. I'll read AI on that for an hour. My sen
18.
▲
by
ddp26
2mo ago
Is it possible GPT-5.6 is not a very aligned model?
19.
▲
by
ddp26
2mo ago
People have been making claims about the commoditization of llms since chatGPT, and they've been wrong every time as quality and prices and differentiation have increased.
20.
▲
by
ddp26
2mo ago
But Scott's point is more: why even have markets? Once you have the superforecasting available on the questions you care about, why do you need to publish it for everyone to also react to?
21.
▲
by
ddp26
2mo ago
Almost by definition, once AI forecasters are in the market, they won't (all) be beating the market. But why evaluate AI forecasters by beating the market? Do we evaluate deep learning by whether hedge funds make money from it in the m
22.
▲
by
ddp26
2mo ago
Doesn't this argument prove too much? Why does AlphaSense sell their company research instead of using it to trade themselves? Why do people work on open source time series forecasting packages instead of quietly using them to trade?
23.
▲
by
ddp26
3mo ago
Indeed they are! It's funny, in Sept 2024 I and others wrote about how the AI Superforecasters _weren't_ here, despite several claims that they were: https://www.lesswrong.com/posts/uGkRcHqatmPkvpGLq/cont
24.
▲
by
ddp26
3mo ago
Yeah, a great developer I know showed me how he could use it to get a safe dev container for Claude Code, in a way that wasn't doable with Docker.
25.
▲
Porting the Moebius 0.2B image model to run in Claude Code on web
(simonwillison.net)
2 points
by
ddp26
3mo ago
|
0 comments
26.
▲
The Wealth of the Richest People in AI
(futuresearch.ai)
4 points
by
ddp26
3mo ago
|
0 comments
27.
▲
by
ddp26
3mo ago
Fair point.
28.
▲
by
ddp26
3mo ago
Is this the trend? There have been various points where one of Anthropic or OpenAI was substantially ahead. Sure, many times they're close, but now doesn't seem like one of them.
29.
▲
by
ddp26
3mo ago
Based on my conjecture that Anthropic is ahead on AI research, and that OpenAI doesn't know how to make Fable-class models.
30.
▲
by
ddp26
3mo ago
I'm going to pre-register my prediction that GPT-5.6 Sol is significantly behind Claude Fable 5, as evaluated by general consensus once time has passed for people to get familiar with both.
More ›