5 ms·
Curious on why you think this. Any data points that led you to this?
by ryeguy_24 5mo ago
Curious on why you think this. Any data points that led you to this?
- howdareme 5mo agoThe benchmarks they released
- johnfn 5mo agoWhat do you mean? In most cases, the benchmarks show a larger number for Muse and a smaller number for Opus.
- spprashant 5mo agoIn Multimodal yes, but Opus is definitely edging out in Text/Reasoning and Agentic benchmarks. I think the general skepticism is because they are late to race, and they are releasing a Opus-4.6-equivalent model now, when Anthropic is teasing Mythos.