Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
trjordan
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
11 ms
·
1.
▲
by
trjordan
13d ago
Grok models are struggling too: https://status.x.ai/ reply Looks like trouble in the SpaceX datacenters.
2.
▲
by
trjordan
16d ago
Related: https://www.birdweather.com/birdnetpi I really gotta set mine up!
3.
▲
by
trjordan
21d ago
An intuitive explanation is that financial products are, approximately, buying and selling as part of the same transaction. You can't separate the "selling premiums" part from the "paying out claims" part. This is t
4.
▲
by
trjordan
26d ago
Not mentioning Grok 4.6 here is a crime. Fast and accurate. And it can communicate, unlike the gobbledygook that comes out of Claude.
5.
▲
by
trjordan
1mo ago
FWIW, this particular header in news articles is a deliberate choice originated by Axios. It’s notable enough and effective enough they wrote a book about it. https://www.axios.com/smart-brevity Did Claude write this articl
6.
▲
by
trjordan
1mo ago
It’s probably worth remembering that system prompts are part of a layered system of shaping Claude’s behavior. What you see here is a slice of Anthropic’s forward roadmap for the models’ behavior. > When a person is in crisis or expressi
7.
▲
by
trjordan
1mo ago
I deeply love this idea of specialized LLMs for search. It's also extremely confusing to me how rough Google's entrance here is. When I, a human, need an answer to anything moderately complex, it's unlikely that I get it on t
8.
▲
by
trjordan
1mo ago
If we define taste as the intuitive act of saying "no, again," then I fully disagree with this whole article. Every AI-frustrated (but LLM-written, sigh) blog post about the loss of taste and craft and hard work in development sou
9.
▲
by
trjordan
1mo ago
I like the idea of devtools being open source. I also like them to work, and that's what I value more that philosophical purity. Take his side project, Meat. I've been chewing on this problem for a while. It's a real problem
10.
▲
by
trjordan
2mo ago
Man, it's wild how I have no original thoughts. I've been pulling on this thought this morning, complete with checking in on how CodeSpeak and Tessl are doing. I'll add this link to the pile: https://martinfowler.c
11.
▲
by
trjordan
2mo ago
> Claude wasn’t just a compiler here. I never handed off a task and let an agent make a bunch of decisions in order to reduce it to practice. > I’d say that, in all the ways that matter, I understand the code. I think the dissonance h
12.
▲
by
trjordan
2mo ago
> Breed: mixed That's a corgi. > Pushinka subsequently became irascible, and "a little nippy" according to Caroline Kennedy, which she attributed to her upbringing in a scientific laboratory. No it's because she&#x
13.
▲
by
trjordan
2mo ago
So, we tried feeding the logs back to the LLM, and it mostly produced slop. Lots of decisions nobody cared about. The biggest things that moved the needle were: - Baseline it. We mine previous logs, github comments, etc. for "what you
14.
▲
by
trjordan
2mo ago
100% important. But what decisions do you care about seeing? The whole point of the agent is to make decisions for you. If you want to make every little detailed decision, just write the code. The whole art of this problem is figuring out w
15.
▲
by
trjordan
2mo ago
> "cleanupPeriodDays": 99999 Throw that in ~/.claude/settings.json
16.
▲
by
trjordan
2mo ago
The agent will always fill in the gaps in your understanding. It's not a compiler. It's categorically different from any of the other ways we've built software. I'm not sure reading code is coming back. The ritual of rea
17.
▲
by
trjordan
2mo ago
I am no fan of Zuck. But this is his whole deal. Instagram was a purchase. Facebook wasn't his idea. Threads is a copy. The 1 thing that Zuck understands better than anybody is that engagement is the only thing that matters to social n
18.
▲
by
trjordan
2mo ago
lmao hi Matt I agree, though maybe the middle ground is something more like: the constraints of our environments shape us. It's easy to say that big companies are a weird and unique cave that produces weird and unique outcomes, but oth
19.
▲
by
trjordan
2mo ago
Most startups fail. Most big company projects are kind of worthless. These are two sides of the same coin. Producing something novel and valuable is HARD. Unbelievably hard. The idea is hard. The building is harder. The scaling and steering
20.
▲
by
trjordan
2mo ago
AI is so miserable for this. It's so focused on doing what you ask, it forgets that there's stuff worth doing that you didn't ask for, like defining reasonable abstractions. Getting away from stuff like this is exactly why I
21.
▲
A compiler that never says No
(tern.sh)
2 points
by
trjordan
2mo ago
|
0 comments
22.
▲
by
trjordan
2mo ago
I was heading to dinner with a friend who worked in infra. Google maps said we could bike across town in 20 minutes. He suggested we leave 40 minutes ahead of time and grab a drink at the bar if we got there early. When I raised an eyebrow,
23.
▲
by
trjordan
3mo ago
It's because it mostly doesn't matter what you are trying to get the code to do. What matters is what the code does. Session logs can absolutely be useful, but not when building further. It's just that that the place they slo
24.
▲
by
trjordan
3mo ago
100%. The problem with them isn't making sure they're doing the right thing, it's making sure they're not making bad assumptions. IMHO this is where code review goes until we fix the individualized model thing: you need
25.
▲
by
trjordan
3mo ago
This is RL, right? Like, this is exactly why models have mostly converged around obvious style, because we train them literally on thumbs-up/thumbs-down data of what good behavior and good code looks like. And that's why it's
26.
▲
by
trjordan
3mo ago
You can't unit test for taste if you haven't written down what you mean by taste. If you can externalize it, then you can. Follow this line of thinking, and the AI-friendly answer is easy: we just have to externalize everything we
27.
▲
by
trjordan
3mo ago
If you didn't take the time to write it, why should I take the time to read it? This is a band-aid. Maybe even a good band-aid, because it'll keep individual contributors from flooring the zone. But the core problem is Github'
28.
▲
by
trjordan
3mo ago
I think there's 2 important, but separate, ideas in this post: - Models are not good at or getting better at creating strong invariants, which his fundamental to good software - It is unclear how to keep tabs on what the agent is doing
29.
▲
by
trjordan
3mo ago
The worst thing that can happen at an early company is that it sort of works. I like the deal where I roll the dice and don't have to work again if I win. I'm fine with the deal where I take a barely-passable salary and do somethi
30.
▲
by
trjordan
3mo ago
The core of the problem is that there are a million tools that make AI better, and no ways to measure whether AI is working better. Big companies with popular products have it. They do something between normal product analytics and chatbot
More ›