Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
extr
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
extr
12d ago
I use NextDNS at home. It's great to have lots of control around DNS while also running a locked-down router (Eero). I personally use it in combination with a raspberry pi to run a version of this https://cazander.ca/20
2.
▲
by
extr
27d ago
Sounds like you have a lot of axes to grind outside of simply "the latest models regressed on delivering concise prose".
3.
▲
by
extr
27d ago
It's "literally unbearable" when the AI that completes software engineering tasks at 100x speed and quality from 2 years ago uses too much jargon?
4.
▲
by
extr
27d ago
I'm sorry but the whining over LLM output styles is embarrassing. Do Claude and GPT models always respond in exactly the way my most articulate coworker would? No. The overused jargon is absolutely annoying. But these things aren'
5.
▲
by
extr
29d ago
Unfortunately Sol does not compare to Fable at all.
6.
▲
by
extr
1mo ago
keep in mind fable = mythos which as been "done" since february. so the gap is not 2 months, it's more like - techniques probably started "working" in late 2025, now are trickling down to 2nd tier labs 9 months late
7.
▲
by
extr
1mo ago
It's because Fable is just synthetic RL tasks + scale. The secret has been out for awhile now.
8.
▲
by
extr
1mo ago
Manus did a lot of harness work to make up for gaps in Opus 4.5 tier models. I tried it for a time - they had a great deep research/PDF generation pipeline, parallelization, etc. The bitter lesson has now come for them: the latest mode
9.
▲
by
extr
1mo ago
yeah it's true, you do have to guide them. i find that the key is you have to know what's possible. you have to have the instinct for "this really shouldn't be so difficult". my junior SWE coworkers have the same tr
10.
▲
by
extr
1mo ago
$80 is definitely low now that I look at my numbers. but not OOMs low, it's closer to like $200 on heavy days. i don't know how you're doing $3k/day, that's wild. i'm pretty aggressive about compaction and sess
11.
▲
by
extr
1mo ago
Yes lol. Of all things people are getting on me for it's the number of LoC x Years In Business of this startup. I don't fucking know, I didn't start the company and I wasn't here for several of those industrious years. L
12.
▲
by
extr
1mo ago
> "unguided" means "I typed a prompt into claude code and waited yolo" Yes, this is literally what that means.
13.
▲
by
extr
1mo ago
This is a great point and I agree. My own productivity varies based on what part of the codebase I'm working on. If it's "been in there before" and I know the right questions to ask, I can one-shot a good design/imp
14.
▲
by
extr
1mo ago
It's a fair point, it's not truly unlimited and I do wonder how that would change my workflow. I can definitely imagine if I was inside Anthropic or OAI with unlimited "fast" tokens, you would be more tempted to hand ove
15.
▲
by
extr
1mo ago
No, actually. The point is to build a profitable business.
16.
▲
by
extr
1mo ago
> unguided LLM usage Why aren't you guiding your LLM usage? Is that what I said - to spam it and not guide anything? Or to have a careful workflow where you agree on design and maximize your human judgement/leverage? > any s
17.
▲
by
extr
1mo ago
How is this not true? Taking a Senior SWE @ ~$200K, even just the base salary cost / 2080 working hours is $100/hr. Fully loaded employer cost + accounting for non-coding time gets you to upper 100s easily. Even for a junior makin
18.
▲
by
extr
1mo ago
Yes 100%. This morning I casually prompted Codex to drive the browser to complete extensive performance testing in-situ that would have literally been weeks of work before. Probably in reality it just wouldn't have been done, and perfo
19.
▲
by
extr
1mo ago
Have you worked at many startups?
20.
▲
by
extr
1mo ago
This was more true a few months ago but Fable has improved the situation considerably. Also just remember - minimalist code looks and feels great but customers do not read your code. I have caught myself many times providing "correctio
21.
▲
by
extr
1mo ago
Disagree. I operate this way inside a multi-million line legacy codebase.
22.
▲
by
extr
1mo ago
Keep the decision-making and execution separate. Use the high IQ models to chat about the design and make them drive subagents to do the actual work. "Chat" style threads are actually quite cheap. Where it gets expensive is having
23.
▲
by
extr
1mo ago
Performance is better than ever. It's never been more practical to set up wildly complex synthetic test environments and measure perf wins. Plus the models will find every possible algorithmic/design improvement. It actually gives
24.
▲
by
extr
1mo ago
I would be really curious to hear from devs at Databricks what the experience of development is like internally. I work at a small startup with essentially unlimited AI spend budget - the entire point is that I should be turning to it at ev
25.
▲
by
extr
2mo ago
I don't really care if it was made by AI or smeared onto the keyboard by a monkey. It was effective in it's job and communicated clearly.
26.
▲
by
extr
2mo ago
Well made website IMO. Gets to the point with plenty of easy-to-understand evidence and calls to action.
27.
▲
by
extr
2mo ago
Seems pretty reasonable.
28.
▲
by
extr
3mo ago
I think DeepSWE is flawed in a different way: the tasks look like someone took a bunch of big highly technical PRs they found really well done, and inverted it into specs for agents to autistically execute. This is not really how people use
29.
▲
by
extr
3mo ago
[Future voice]
30.
▲
by
extr
3mo ago
I'm skeptical any of that matters at all if at some point AI is perceived by the government to be a true existential risk to public welfare.
More ›