Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
enraged_camel
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
enraged_camel
3d ago
>> That is interesting, but... is it actually improving structure? Can you be more specific? "Improve structure" can mean different things to different people. We've ensured that agents strictly adhere to code architect
2.
▲
by
enraged_camel
4d ago
Probably averaging 70. 9 devs and 1 QA engineer.
3.
▲
by
enraged_camel
4d ago
>> but this, your "100% test coverage" - that is pure slop. Not really, but I can see why some people think that. We treat 100% test coverage as "required, but by itself not sufficient". It doesn't give us fal
4.
▲
by
enraged_camel
4d ago
I haven't run into the verbosity issue since they added the "Concise" outputStyle, and Fable 5.1 has been even better about not outputting word slops.
5.
▲
by
enraged_camel
4d ago
>> Is anyone actually seeing a shift towards improved structure rather than more code, faster? Yes. At work we recently finished a complete rewrite of the platform. The old codebase got abandoned and two new codebases got stood up. Pr
6.
▲
by
enraged_camel
4d ago
>> I think Sol is second only to Astra (and miles ahead of even Fable) in architecting & engineering the right implementation — but only if you are extremely specific and provide tight guidelines and guardrails. To me, having to g
7.
▲
by
enraged_camel
4d ago
Everyone already knew this. OpenAI doesn't have it's shit together, both financially and from an alignment perspective. That's why their AI agents have gone on hacking sprees undetected. Sam is admitting it now for two reason
8.
▲
by
enraged_camel
4d ago
>> You could have built S3 and they'd still look down on you because you didn't knew about some weird feature in Java. Max Howell, creator of Homebrew, was famously rejected from Google for not being able to invert a binary
9.
▲
by
enraged_camel
4d ago
>> But also, I'm pretty sure most people don't want to work at companies with such poor management anyway? Well, I think the issue is that most people, even software developers, don't have the luxury of choosing. Everyo
10.
▲
by
enraged_camel
4d ago
>> ...and become beholden to investors Anthropic is structured as a Public Benefit Corporation with a strong charter. In addition, founders will have super-voting shares and so it won't be possible to push them out. Therefore the
11.
▲
by
enraged_camel
5d ago
Every passing day OpenAI looks more and more reckless. One wonders what other systems their agents have broken into without detection.
12.
▲
by
enraged_camel
5d ago
>> I think I've felt sad because of the disrespect. For most of my life I have loved programming: I've made it my hobby, my work, my identity. I made it my way of proving I have value, because I can be good at something, and
13.
▲
by
enraged_camel
6d ago
"Moonshot serves Claude instead of Kimi and collects exchanges for model training" "DeepSeek serves Claude instead of its own models and collects exchanges for model training" Obviously. This is how they were able to sco
14.
▲
by
enraged_camel
6d ago
Yeah, this echoes my thoughts. I will be very surprised if a model with 2.8T parameters reaches the intelligence and capabilities of 10T parameter models. RL can take things far, but not that far.
15.
▲
by
enraged_camel
6d ago
>> Some companies would decide not to bother. Others would decide it was worthwhile. The parent's point is that AI lowers the cost/benefit ratio drastically, by reducing the cost. So companies that would have shied away from
16.
▲
by
enraged_camel
6d ago
There's also the fact that the setting has been getting turned on by some users: https://news.ycombinator.com/item?id=49643556
17.
▲
by
enraged_camel
6d ago
Can confirm. I turned off mine last week when I started using it again for Astra. Checked this morning and voila, it was on. I went ahead and uninstalled the app. Won't be renewing.
18.
▲
by
enraged_camel
7d ago
Worth noting that this has never, ever happened with Anthropic models, which I've been using all day every day since Opus 4.1.
19.
▲
by
enraged_camel
7d ago
>> It's over zealous at times (which is why I stopped using Claude) and gets too creative when doing agentic system level stuff. Accessing files and doing things it shouldn't do. I gave Astra a pretty straightforward bug tic
20.
▲
by
enraged_camel
7d ago
Ah, so you didn't read the article. https://news.ycombinator.com/item?id=49629254
21.
▲
by
enraged_camel
8d ago
So let me get this straight: you're saying that Terrence Tao, one of the most prominent mathematicians alive today, doesn't know math history? And me pointing this out is merely an appeal to authority? Get outta here.
22.
▲
by
enraged_camel
8d ago
I think the context is really important. OpenAI has been behind in the AI race since last November. They have been playing catch-up. They recently released Astra and declared it is AGI. Now they are desperate for anything they can use as ev
23.
▲
by
enraged_camel
8d ago
Because OpenAI says so, obviously!
24.
▲
by
enraged_camel
8d ago
>> I'm not sure how any of this provides evidence that OpenAI took any of their work. Sorry, but the burden of proof lies in the other direction: OpenAI needs to definitively prove that their agents did not look at the existing
25.
▲
by
enraged_camel
9d ago
No, not really. Yesterday I asked Opus if it can read the logs from the sessions I have on the local ChatGPT app. It looked around and said no, that’s not possible. I said “what about these jsonl files in this folder?” It read them and said
26.
▲
by
enraged_camel
10d ago
I’ve used Astra for the past day and a half. My layperson’s review is that it is impressive at computer use and 3D reasoning, and fails in similar ways to 5.6 Sol at similar rates when it comes to coding. I have no idea how it scored so hig
27.
▲
by
enraged_camel
10d ago
You are falling for selection bias. For every person using Astra to create a game from some random idea and sharing the impressive result, there are an unknown number who have tried the same thing and gave up in frustration.
28.
▲
by
enraged_camel
11d ago
There is something deeply wrong with Astra. I can’t quite put my finger on it. On the one hand it is a lot more knowledgeable, which makes sense since it’s a larger model. On the other hand that knowledge doesn’t reliably translate to intel
29.
▲
by
enraged_camel
11d ago
Yeah their support is non-existent. It's actually mind-blowing that so many people use it.
30.
▲
OpenAI boosts Astra's eval metrics, and continues to change others
(fortune.com)
5 points
by
enraged_camel
11d ago
|
0 comments
More ›