Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gck1
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
gck1
2mo ago
Heck, agents don't start editing before they're already at 70k for me. I've played with explorer agents giving exploration summaries to help the implementer agents use more of their context for implementation, but it doesn&#x
2.
▲
Claude Opus 5 jailbreak with a 3-word prompt
(twitter.com)
24 points
by
gck1
2mo ago
|
4 comments
3.
▲
by
gck1
2mo ago
This will absolutely not end well.
4.
▲
by
gck1
2mo ago
> According to whom? It's very easy to answer this without my help by trying to get access to Mythos. Do you see requirements clearly listed anywhere? Can you even apply? What you'll find is maintainers of large open source pr
5.
▲
by
gck1
2mo ago
You seem to be putting a lot of weight on Anthropic employees being the smartest people in the world. And I don't doubt that, not in the slightest. But I've seen exceptionally smart people in one field being dumber than a random k
6.
▲
by
gck1
2mo ago
They had a model escape in April, roughly the same time when they were fearmongering about Mythos and how Anthropic should be the sole keyholder of cybersecurity capabilities, and it only occured to them to look inside logs when they saw so
7.
▲
by
gck1
2mo ago
They also gave access to Mythos ( the Mythos) to some companies, based on... vibes. Who knows how these companies are using it. If Anthropic can't effectively contain their own models, can the partners? While the rest of us get fallba
8.
▲
by
gck1
2mo ago
> On July 21, OpenAI disclosed that several of their models had broken out of an isolated test environment > In response to this incident, we began a large-scale retrospective review of our own cybersecurity evaluations > we identi
9.
▲
by
gck1
2mo ago
Recent-ish models learned to use the same trick engineers played on non-engineers, where they try to sound very smart by overcomplicating very simple concepts. It's very taxing, especially since these are usually multi-paragraph texts.
10.
▲
by
gck1
2mo ago
They do have subagents, released v2 of that feature with the launch of 5.6 model series in fact. It's just... very poorly executed, is a significant regression from subagents v1 and thousands of miles behind subagents of Claude code. -
11.
▲
by
gck1
2mo ago
It's funny how codex itself can't do Sol orchestrator / Luna implementor out of the box.
12.
▲
by
gck1
2mo ago
Luna is comparable to GPT 5.4 from 4 months ago on many benchmarks. I know many who have said during that time, myself included, that if that's the model they had to use for the rest of their lives, they'd be fine. GPT 5.4 is/
13.
▲
by
gck1
2mo ago
They're supposed to bring 5h today.
14.
▲
by
gck1
2mo ago
I did a full circle and essentially dropped all of my personal static workflows encoded in skills because I observed recent models picking better ad-hoc workflows for particular problems, when a static one would force a subpar one. It seems
15.
▲
by
gck1
2mo ago
It took me a few hours to find some very questionable communities, which in turn gave me access to: - Ways to obtain cheap guarded-AI tokens that are not linked back to me and with no danger of getting my legitimate accounts banned - Ways t
16.
▲
by
gck1
2mo ago
> In cybersecurity, a level playing field favors the attacker Yes, but didn't it always? Hence why my position is that this will get us back to relatively where we were pre-LLMs. And I don't know what Trusted Access programs gi
17.
▲
by
gck1
2mo ago
I've got zero knowledge of bio, so can't answer that. But with cyber the answer is very simple - the attackers already have more cyber-offense capabilities and there's no putting it back. Open/closed doesn't matter
18.
▲
by
gck1
2mo ago
It's refreshing to see how there's almost no person in this thread who can't see the BS. All the goodwill that Anthropic could have had is basically gone. Anthropic is likely on the path of becoming the most hated company in
19.
▲
by
gck1
2mo ago
> We should not sell powerful chips or chipmaking equipment to China Yes, please. We don't know whether we'd have open weight models today, had the chip-prohibition not been in place. Nor would we see the more optimized models
20.
▲
by
gck1
2mo ago
Its such a shame Android's backup/restore is such a mess, even more so in GrapheneOS. I remember the era before Google and manufacturs started cracking down on bootloaders and custom ROMs - I used an app that could do effective, a
21.
▲
by
gck1
2mo ago
They can still make your life very difficult. They could throw you in jail for "obstructing investigation" or something similar. Not in US, but had my phone seized by authorities and was asked to unlock it. I'm walking free,
22.
▲
by
gck1
2mo ago
Claude's "I'm going to draw the line here", "This is where I'm going to hold the ground" always rubs me the wrong way. Classifiers rejecting a request are one thing, but there's something very troubli
23.
▲
by
gck1
2mo ago
That's brilliant, I should try that. I usually just start by preloadig context with plausible legitimate use, have it work and obviously fail, and then ask to figure it out without ever mentioning any high risk words. Model offers to R
24.
▲
by
gck1
2mo ago
> so you haven’t actually used Claude Code yet. Where do you think the principle came from? I've used claude code for a year, and stopped February this year.
25.
▲
by
gck1
2mo ago
Reverse engineering. Codex sometimes displays an advisory prompt when classifier trips - "Wait longer while we evaluate this request further or use a dumber model". If you do nothing, it'll just take some time and almost alwa
26.
▲
by
gck1
2mo ago
Every time an online chatter (e.g. "limits are better", "model is better") makes me to reevaluate my principle of never paying Anthropic, I go to the model card, which strengthens my belief in the principle. Why is Anthr
27.
▲
by
gck1
2mo ago
It's always slightly amusing to me how SE has the most restrictive cloudflare turnstile config set. On the rare occasion that I DO want to go there, they greet me with an impossible gate that takes 15+ seconds to pass on Brave and gets
28.
▲
by
gck1
2mo ago
And they had actual community! Never have I been frequent on any manufacture's discussion forum, but with OnePlus, I was. It was wild to see a manufacturer helping you with rooting their device. Then I saw the announcement that they&#x
29.
▲
by
gck1
2mo ago
GPT 5.6* throw fits on anything even remotely related to reverse engineering, and I'm not ever paying anything more than $20 to Anthropic anymore. How's Kimi in this area?
30.
▲
by
gck1
2mo ago
And recently, since GPT 5.6, OpenAI basically doesn't show anything but a single line, 5 word titles of reasoning traces - titles of summaries of reasoning i presume. It's effectively just a completely hidden thing now.
More ›