6 ms·
Claude: Elevated Error Rates for Opus 4.8, Opus 4.7, Opus 4.6, and Sonnet 4.6
- bombcar 3mo agoAPI Error: 529 Overloaded. This is a server-side issue, usually temporary — try again in a moment. If it persists, check https://status.claude.com https://status.claude.com.
- bombcar 3mo agowe may be coming back online
- jesse_dot_id 3mo agoI noticed it in real time, unfortunately. Perhaps the token bonfire I've been feeding all day is to blame...
- NoHatH4k3r 3mo agoThey have been screwing around in the back-end, for a while. They are doing something and not being transparent, re-routing the models, degrading performance and intelligence and all week post-fable ban the quality was garbage, last day and a half it improved, last half an hour degraded, now it's down.
- dshat 3mo agoSeen this yesterday as well.
- dannyw 3mo agoInterestingly not Opus 4.5 apparently, which is still available via API.
- danielrmay 3mo agoPerhaps Claude approved the wrong compute allocation plan
- bombcar 3mo agoHow can you switch vsclaude to opus 4.5?
- minimaxir 3mo agoBetter buy more capacity from SpaceX.
- behnamoh 3mo agoTimes like this remind me that despite GLM and Codex and other models being hyped up as Claude Opus 4.8 replacements, I still would not trust them with my most important work. For example, right now I'm working on a huge refactoring project, and even Opus has struggled with it after several days. I cannot even imagine how GLM, Codex, or other models would handle this. So the only option for me is to wait until this outage is over. And it's not like open models are cheap to run even as alternatives. For example, with my $100/mo subscription for Claude Code, I often burn more than $100 a day several times a week. But if I were to use the API of GLM, it would be about $300.
- lkois 3mo agoSince you cannot imagine how they'd perform, isn't this the perfect opportunity to test your assumption?
- mixtureoftakes 3mo agoyeah and lets not forget codex and glm have subscriptions too, with even more usage per dollar
- behnamoh 3mo agobut they also burn more tokens per task, so in the end, Claude comes out as the more efficient one, despite giving you less tokens.
- viraptor 3mo agoYou've got it backwards. Opus is the token/money burning one https://deepswe.datacurve.ai/ https://deepswe.datacurve.ai/ Gpt 5.5 uses a third of the opus 4.8 tokens for the same task and scores higher. Glm 5.2 was worse in quality but used half the tokens - 5.3 is not tested yet but will be higher.
- Der_Einzige 3mo agoA lot of the perception of open source models being garbage is that they're still using the same piss-poor sampling algorithms that OpenAI/Anthropic force on their users, i.e. Top-p, top-k. These lead to small accumulation of sampling errors which makes it all but inevitable that open source models will shit the bed by the 200K token mark or even sooner. If you set your opencode to use a good sampling algorithm, such as min_p or top-n sigma (llamacpp supports both), you'll find that at least for long running tasks, your model gets a lot better. It won't make GLM as good as Opus 4.8, but it will stop the feeling of "brain damage" from running open source models at the edge of their context windows. And yes, there is an upcoming (hopefully NeurIPS) paper titled "Long Context Generation is a Sampling Problem" for more details about this. Give it two months and it'll be on Arxiv one way or another.
- hmokiguess 3mo agoIt has been like that all of last week
- kinduff 3mo ago529s as a forcing function for taking a walk
- internet2000 3mo agoFix Fable pls.
- airstrike 3mo agoThis happens all the time. Nothing to see here
- cadamsdotcom 3mo agodang et al, "Service X is down" is not "news". Kinda feels like points-farming. Anyone caring that X works is gonna know it's down - they can't work! And probably why they are at HN ;) HN as an "is service X down?" detector is less reliable than trying the actual service :) If HN is gonna keep letting people post "X is down" no worries. But it seems worth a flag.
- sb057 3mo agoPer their status page, the main product now has one 9 of uptime.
- jasonjmcghee 3mo agoSonnet 5 here we come
- nisten 3mo agoMy speculation is that we'll get fable back tomorrow. Usually we see messy devops stuff or maybe nerfing(not much these days) around the time they're releasing new models
- ChicagoDave 3mo agoI have two sessions going. It’s my fault. Sorry guys.
- cfbradford 3mo agoRookie numbers
- shinobi-apps 3mo agoThis mess sounds like a begining to skynet:)
- jmward01 3mo agoLots of downtime on CC the last few days. They also pushed an bad release to CC that kept doing 'No response from API · Retrying in .... check your network'. 'claude install stable' fixed that. How I was not on stable I have no idea. I can only guess how many tokens I was billed for sending requests that CC got 'No response from API' for. I have a feeling the big AI vendors won't have the loyalty that IDEs and other dev tools generate. They really need to work on trust now to avoid people hopping later.
- swader999 3mo agoMaybe fable coming back
- ninja3333 3mo ago[dead]
- deleted 3mo ago[deleted]
- deminature 3mo agoOpus 4.8 is nearly unusable at the moment. Making extremely obvious errors, failing to debug issues even after being prompted with the exact line of code where the problem is. It should be illegal not to disclose serious degradations of performance while still charging full price. It may just be my region's datacentres degraded or something. The company is so opaque about what's happening, it's impossible to know what is actually going on. The optimistic take is they're preparing to relaunch Fable 5 in the next day or so and siphoning off power for that. Perhaps they're doing a snap retraining to strip out any cybersecurity capability.
- naveedahmedswe 3mo ago[flagged]