Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gaflo
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
gaflo
1mo ago
You are displaying ads on your website, there's a clear financial incentive.
2.
▲
by
gaflo
2mo ago
Which work of his do you respect?
3.
▲
by
gaflo
3mo ago
You're probably right, I guess I'm too optimistic.
4.
▲
by
gaflo
3mo ago
Consider adding a rule that an author must disclose (in their own words) for what parts and to what extent LLMs have been used to assist their project.
5.
▲
by
gaflo
3mo ago
PRNG is deterministic.
6.
▲
by
gaflo
3mo ago
Very interesting, thanks for your reply. I don't have much domain knowledge with your sorts of project so I don't have an intuition on how well it would go with coding agents. I have done quite a lot of programming with graphs, bu
7.
▲
by
gaflo
3mo ago
Can I ask what exactly you are building? Your experience tracks for me when building a real product -- something I want other people to use. Most of my time on these projects is spent talking to my users and carefully refining my requiremen
8.
▲
by
gaflo
3mo ago
Cynical take: If their model was so groundbreaking they wouldn't have to involve the government for their marketing campaign. You would notice shit breaking everywhere; oh wait, how many days has it been since the last supply chain att
9.
▲
by
gaflo
3mo ago
From how Simon described it it's not a native feature, but one that the model built as a solution for automatically testing. You could already instruct the agent to write a program that saves screenshots to disk and then reads it. As l
10.
▲
by
gaflo
3mo ago
Thanks for documenting your personal observations. I do have a few questions. First, could you expand by giving other examples on how you observed this model to be relentlessly proactive? From my personal experience with prior frontier mode
11.
▲
by
gaflo
4mo ago
Is there any credible primary source for this exploit being real?
12.
▲
by
gaflo
4mo ago
If you upgrade your 8 year old phone the many incremental upgrades will be very noticeable. From my personal experience the LLM space is also moving at a faster pace than the phone industry at the moment, but at least from a financial persp
13.
▲
by
gaflo
4mo ago
What kind of data are you interpreting? Do you mean document extraction from different languages? I have only used GPT5.5 for agentic coding, which did get significantly better from my experience, although that does align with your conjectu
14.
▲
by
gaflo
4mo ago
Can you elaborate what kind of system you built? I'm curious what specific prompts are getting worse responses with the newer models.