7 ms·
(I work at Anthropic) We have publicly stated[1] that our goal is to deploy Mythos-class models at scale when we have the requisite safeguards for offensive cyb
by smca 4mo ago
(I work at Anthropic) We have publicly stated[1] that our goal is to deploy Mythos-class models at scale when we have the requisite safeguards for offensive cyber risks in place. Mythos is a general frontier model, not a cyber-specific model so there are many reasons why we think our users will benefit from access (with the aforementioned safeguards in place) in due course. Compute has also not factored into our decision[2] to rollout the model in a limited fashion to defenders. We'll be sharing more soon on the first month or so of the project and rollout.
[1] https://www.anthropic.com/glasswing#:~:text=deploy%20Mythos%2Dclass%20models%20at%20scale https://www.anthropic.com/glasswing#:~:text=deploy%20Mythos%...
[2] https://x.com/logangraham/status/2054613618168082935 https://x.com/logangraham/status/2054613618168082935
- yanis_t 4mo agoAre there any publicly verifiable sources that Mythos is that much more intelligent than Opus, so to be considered much more dangerous (as it is presented in the public discourse by Anthropic)
- empath75 4mo agoIt doesn't have to be _much more intelligent_ than Opus to be a risk. It doesn't even need to be _more intelligent_. It just needs to be _better at finding security problems_. Which could happen from just minor improvements in training data, or the harness, etc. Even a small improvement could shift it from finding very few new security holes, to reliably finding many at scale.
- enraged_camel 4mo agoYeah, I think a lot of the disconnect here is that people think of "model intelligence" as some sort of IQ score, rather than a combination of scores that measure abilities at a large variety of domains.
- temac 4mo agoWeird take to claim "generally intelligent frontier" (whatever rhat means) and restrict availability based on "offensive" cyber security alone (how can this be handled at all compared to fixing software also remain to be seen) all while competitors but more importantly sw maintainers (eg curl) estimate that the capability in finding cybersecurity bugs is similar to what other modern models produce, and this has just significatively risen in the last months for everybody.
- alt227 4mo agoMultiple people who have already used Mythos or been given its reports on their software have publicly stated that it's all hype, and that it is not really finding any new critical bugs which other models cant.
- Lerc 4mo agoDo you have any good sources on that? I have seen things to suggest that not all of the hype is true, but so far I have not encountered anyone claiming all of the hype is untrue. Which is what I interpret "its all hype" (sic) to mean.
- alt227 4mo agoFor example, It was recently let loose on cURL and its maintainer is less than impressed: https://www.theregister.com/security/2026/05/11/anthropics-bug-hunting-mythos-was-greatest-marketing-stunt-ever-says-curl-creator/5238111 https://www.theregister.com/security/2026/05/11/anthropics-b...
- Lerc 4mo agoIf you remove the fluff that the register added and stick with https://daniel.haxx.se/blog/2026/05/11/mythos-finds-a-curl-vulnerability/ https://daniel.haxx.se/blog/2026/05/11/mythos-finds-a-curl-v... it seems like a claim that's a claim fairly distant to "its all hype". Less than expected perhaps? Maybe the code really is unexpectedly robust? I guess time will tell on that point.
- alt227 4mo agoHis direct quotes are: > "My personal conclusion can however not end up with anything else than that the big hype around this model so far was primarily marketing." > "I see no evidence that this setup finds issues to any particular higher or more advanced degree than the other tools have done before Mythos." > "An amazingly successful marketing stunt for sure." Personally I see this as a very strong claim of hype. I take away from this that Mythos is a hyped up marketing stunt, and not what it was preesented to be by Anthropic at all.
- grayhatter 4mo agoare you able to detail a single safeguard you plan to implement, so that I can stop believing it's vaporware and/or a scam?
- GenerWork 4mo agoHow would it be vaporware? It's out in the wild and has been used by individuals/corporations.
- grayhatter 4mo agoThe security/safety controls they have to add to make it safe enough to release?
- saidnooneever 4mo agono risk is added. all risk is already maxed out. release it.