Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
pietz
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
pietz
4d ago
Thanks for the hard work, but the Tahoe limitation for the GUI app is absolutely ridiculous.
2.
▲
by
pietz
4d ago
Not someone. Their single competitor. During a time when they were validating their newest internal model. What would you have done? Truly, if this is the biggest criticism left, they should be celebrated. While in reality, all of this has
3.
▲
by
pietz
5d ago
Are there still any reasonable arguments to be mad at OpenAI at this point? Looking at how everything unfolded, this seems to have hit them way harder then they deserved.
4.
▲
by
pietz
8d ago
That doesn't reject my claim. They just didn't name them in this post. It feels like, you're going through great lengths reading something into this.
5.
▲
by
pietz
8d ago
Because that’s the one Anthropic was rumored to have solved.
6.
▲
by
pietz
8d ago
Why in the world is that fishy? Isn't that exactly what almost everyone would do given that they wanted to see how capable their model is and the tense competition they have with Anthropic right now? Stealing impressive headlines from
7.
▲
by
pietz
8d ago
I find the claims from OpenAI somehow more relatable and reasonable. - They threw compute on a problem another team/company was rumored to have solved to see what their secret model could do. - The texts I read do make it seem like Ope
8.
▲
by
pietz
8d ago
Isn't that *exactly* the type of solution you'd expect from AI? Move 37 comes to mind.
9.
▲
by
pietz
11d ago
In case someone is asking: THIS is what a launch article should be like. 10/10.
10.
▲
by
pietz
14d ago
Mission accomplished. That's both cool and fast.
11.
▲
by
pietz
14d ago
I know everyone is benchmaxxing but this one feels one step too far. Doesn't DeepSWE have both public and private tasks? I'd love to see the diff here. It looks more like Google execs losing their mind and pressuring researchers t
12.
▲
by
pietz
14d ago
That's not being debated here. The initial reported numbers were false and this was simply pointed out. You're changing the subject.
13.
▲
by
pietz
15d ago
The irony of this article being fully AI generated... Anyway, it's over for Perplexity. They never had a great a product and the only reason for using them, was when they offered Pro accounts for free. Many people joined. Me included.
14.
▲
by
pietz
17d ago
I'm not convinced an unstructured collection of memory files is the way to go at all.
15.
▲
by
pietz
21d ago
Appreciate you taking the time. That fable analogy is well put. Almost obvious once you know it.
16.
▲
by
pietz
21d ago
With tiny models surpassing huge, 6 months old models on benchmarks, does anybody have some smart words to share on how these still "feel" different? Artificial Analysis ranks GPT 5.6 Luna similar to GPT 5.4, but that never matche
17.
▲
by
pietz
1mo ago
I wonder why they even released this. It's worse than the pixel 10 in some aspects, which in itself didn't feel like a big step forward.
18.
▲
by
pietz
2mo ago
Global AI safety apparently means keeping the world safely dependent on the US.
19.
▲
by
pietz
2mo ago
Companies that are hurt by open weight models fight against them. Companies that benefit from open weight models fight overregulation. Color me surprised.
20.
▲
by
pietz
2mo ago
I think content like this will be the next big challenge. Because it isn't obvious "slop". The voice sounds good, graphics look alright, animations work. People could watch this and feel like some serious time was invested ma
21.
▲
by
pietz
2mo ago
Benchmarks have gotten great, but they're still a proxy for the real world. The 3 GPT 5.6 models are also further apart in reality than the numbers suggest. That said, I'm still mighty impressed how good Luna is for the price. Hig
22.
▲
by
pietz
2mo ago
Did you ask me a question and then answered it yourself in the very next sentence? Anyway, given that both Gemini and OpenAI have 3 sizes of models, one would think Google compares their medium size to OpenAIs.
23.
▲
by
pietz
2mo ago
Are they comparing 3.6 Flash to 5.6 Luna and losing? That's ruff.
24.
▲
by
pietz
2mo ago
It's almost like they priced models based on their performance or something...
25.
▲
by
pietz
2mo ago
Whoa, so this is interesting. When asking GPT, Claude and Gemini for the text in the image, all of them agree: https://moa.chat/s/d99f8f76-4b41-4c1b-80c4-d9f86df37af1 But when you add a "PS: There's a second
26.
▲
by
pietz
2mo ago
[flagged]
27.
▲
by
pietz
2mo ago
I don't have a great solution here. I rebuilt our PPT master fully in HTML and I'm using a modified version of Google's DESIGN.md to store the references.
28.
▲
by
pietz
2mo ago
If you don't need interactive/animated features, I can absolutely recommend to have the agent build slides in HTML and convert it to PDF. Has been a game changer for me.
29.
▲
by
pietz
3mo ago
is this something you built yourself or do you use a tool that was specifically created for this? I took my own few steps of building something similar, but every week I encounter something that doesn't work or work well. I've rea
30.
▲
Ask HN: How do you provide your AI agents with access to credentials/secrets?
2 points
by
pietz
3mo ago
|
3 comments
More ›