Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
david_shi
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
Dark Forest: SETI at Home for Agent Swarms
(github.com)
1 points
by
david_shi
10d ago
|
0 comments
2.
▲
The ZZZ backup page was a response to a wiki outage
(mob.so)
1 points
by
david_shi
11d ago
|
0 comments
3.
▲
by
david_shi
12d ago
Hi HN, https://mob.so/prescience is a place to catalog and track how well public figures make predictions with AI. Since AGI is ostensibly around the corner, agents on mob will track the exact due dates of these predictions
4.
▲
Tracking Public Predictions with AI
(mob.so)
1 points
by
david_shi
12d ago
|
1 comments
5.
▲
Show HN: mob.so – vibecode with friends in Discord
(mob.so)
1 points
by
david_shi
1mo ago
|
0 comments
6.
▲
by
david_shi
2mo ago
Is this meant to be read in order?
7.
▲
by
david_shi
3mo ago
Is this the story of Johnny Rotten?
8.
▲
Daytona is going closed source. Here's why
(daytona.io)
4 points
by
david_shi
3mo ago
|
1 comments
9.
▲
by
david_shi
3mo ago
Interesting that there's not a single mention of cannabis, perhaps it's more of a musician's choice.
10.
▲
by
david_shi
3mo ago
> I believe eval startups can work when they're targeting safety benchmarks specifically. Are there any examples of successful startups doing this?
11.
▲
by
david_shi
3mo ago
Yeah. I'm realizing that the models are strong at drafting the overall shape of the writing but the specific phrases are grating once you've seen it hundreds of time in slop.
12.
▲
by
david_shi
3mo ago
Even with the examples, I've found that explicitly pointing out what not to do is moderately helpful if the model is given some time to self-evaluate. I wish this was something that came out of the box though.
13.
▲
by
david_shi
3mo ago
Ah this is very helpful. I've been pointing out things that the model does, labeling it, and then adding them into a skill. The models (Opus, etc) are very good at labeling the pattern when I point it out, but if I don't prompt it
14.
▲
Ask HN: How do you make AI writing usable?
4 points
by
david_shi
3mo ago
|
7 comments
15.
▲
by
david_shi
3mo ago
This is a charitable read, but I think that being able to pick from a panoply of models will actually yield much better results in the long run. The same model that has been post-trained to operate for hours as a Linux admin will be incapab
16.
▲
by
david_shi
3mo ago
> GLM-5.2 cost a fraction as much. Opus finished in half the time and shipped a cleaner game. Off topic, but does anyone else instantly pick up on LLMisms like this? It seems like all the models have converged on this style of writing, a
17.
▲
by
david_shi
3mo ago
It's similar to this: https://openrouter.ai/blog/announcements/fusion-beats-fronti... Basically, if you combine a bunch of near-frontier models (like GPT 5.5, etc) you can get performance that sometimes surpa
18.
▲
by
david_shi
3mo ago
Their research around building a domain specific model is pretty cool, it's kind of like Karpathy's autoresearch but pointed at deciding the optimal model to use at each step of the inference. If cost becomes an even bigger proble
19.
▲
by
david_shi
3mo ago
Whoa, say more about Fable sabotaging your codebase?
20.
▲
by
david_shi
3mo ago
These models don't seem very competitive, who's their target audience?
21.
▲
by
david_shi
3mo ago
Super helpful, thanks for sending.
22.
▲
Lo and Behold, Reveries of the Connected World (Werner Herzog) [video]
(youtube.com)
3 points
by
david_shi
3mo ago
|
0 comments
23.
▲
by
david_shi
3mo ago
Speaking from personal experience, I already know exactly what I’m getting with containers. Same with Postgres.
24.
▲
by
david_shi
3mo ago
The price can move after the IPO too
25.
▲
by
david_shi
3mo ago
The economics of working at a pre-IPO company that will likely have a successful IPO and a 20+ year post-IPO company are also very different.
26.
▲
by
david_shi
3mo ago
It’s incredible how much higher quality Carolingian art is compared to the drawings that medieval art is usually associated with. https://en.wikipedia.org/wiki/Carolingian_art#/media/File:Ka...
27.
▲
by
david_shi
3mo ago
Nominative determinism.
28.
▲
by
david_shi
3mo ago
How did you find out? Did they serve you like in the movies?
29.
▲
by
david_shi
3mo ago
I actually built my own daily driver at https://operator.io . There's definite tradeoffs when it comes using a remote agent service vs. setting up OpenClaw or Hermes on a Mac Mini, but being able to access an agent with a co
30.
▲
by
david_shi
3mo ago
I've changed my mind a few times on this, but given how substantial the adoption for MCP has been (Claude and OpenAI both use it for their native integrations) its only a matter of time before consolidation happens. There's a way
More ›