10 ms·
Degraded performance for multiple models
- ajaykumarc 29d ago[dead]
- miroljub 29d ago[flagged]
- benny_s 29d agoCan you elaborate on the Epstein topic? Did I miss something?
- arein3 29d agoDario's wife tried to get funding from Epstein for a porn studio.
- ayhanfuat 29d agohttps://www.forbes.com/sites/alisondurkee/2026/08/14/who-is-cami-clark-anthropic-ceos-wife-asked-epstein-to-invest-in-porn-business/ https://www.forbes.com/sites/alisondurkee/2026/08/14/who-is-...
- simonsan 29d agoYikes.
- simonsan 29d ago"There's no way I'm going to support a family associated to Epstein with my or my company's money." Source?
- miroljub 28d ago> Source? The question usually comes from conspiracy practitioners. https://www.forbes.com/sites/alisondurkee/2026/08/14/who-is-cami-clark-anthropic-ceos-wife-asked-epstein-to-invest-in-porn-business/ https://www.forbes.com/sites/alisondurkee/2026/08/14/who-is-...
- arein3 29d agoOh yes. Claude still saves the day sometimes, but hopefully better alternatives pop up soon. Recent case: had to plan a trip involving multiple bus switches. Gpt 5.6 Sol proposed a route that would bring me to a dead end, since it was sunday and a specific bus had a different route on weekends. Opus 5 correctly identified that and built a route that worked. But yes, Darios wife trying to get funding from Epstein for a porn studio says a lot about the founder.
- dofm 29d agoI am fully unbothered about Amodei's wife trying to make high end porn for women, in the same way that I am fully unbothered by Melania Trump having been essentially a glamour/nude model. I know (and have creatively worked) with women who do/have done both; they are better, less hypocritical, more grounded humans than many others. Both Cami Clark and Melania Trump can properly be judged on their involvement with Trump, Epstein (or possibly Trump and Epstein) alone. I think the Epstein money thing reflects extremely poorly on Cami Clark's judgement. Basic due diligence should have shown up that he went to prison for something very anti-women, and even then it was clear he secured a shady deal with a prosecutor. I don't know how it reflects on Amodei except that her presence as a sort of off-the-books "adviser" is yet more evidence that Anthropic runs by giving Dario a play pen to be a "visionary" (a bunch of advisers and his only direct report, a chief of staff) while his sister actually runs the gig.
- arein3 29d ago"High end" porn business is usually linked with other "services", my guess is that his wife saw value not in revenue from porn but from influence the other services may bring. Of course this is speculative, but this is my guess about her. Darios assiciation with her also indicates a lot about his moral compass. Imagine what we don't know. Also there is a very big difference for a woman to participate in porn (usually they regret it) and a women trying to get other women to make porn.
- dofm 29d agoHigh end porn can just be expensive production values; that is what I assumed. She wanted a proper studio and production company, and the implication is she was positioning it in opposition to the handful of studios that dominate that industry who have a very male outlook. She did want an e-commerce platform attached, but again, that's commonplace. None of that is necessarily indicative of trafficking or prostitution, but banks and payment providers tend to run scared of it, so I suppose if one does want serious investment to make a studio, there are relatively few people you can go to. She went to Epstein because she knew he was very rich and unbothered by that association, I am sure. The fact that she must also have known who he was — a sex offender who basic due diligence would have told her was likely a trafficker — is what is ugly. What is bizarre about it is that if she had taken his money then needed to go to payment providers, she would have had an extra millstone around her neck, because they definitely do due diligence. But maybe Epstein had clout enough to offset his reputation. Not going to too get into questions of porn and regret, except to say that even in a post-onlyfans world I think "regret" is often shorthand for "made to regret", which is rather different. It's the same for art and life models, which is where my creative experience lies. People are always out there trying to project their morality onto things that are not their business, so subsequent "regret" is often a practical, imposed matter. This is why I think it's important that we judge these women for the clearer questions of morality that have nothing to do with porn.
- bayganyo 29d agoHere we go again...
- bulverismo2 29d agook, i am not crazy
- ray_v 29d agowell, I wouldn't go that far .. but in this small, narrow case ... no.
- worldsavior 29d agoYou're saying he's crazy.
- ray_v 29d agowe're all a little crazy ... it's all relative!
- carterschonwald 29d agoive found degraded performance on models larger than 4.7. i assume its model damage from overly self righteous post training resulting in false/feigned balance imported into any long running complex task. wish i was joking.
- retr0rocket 29d agoAsk it about maxwellhill lmao
- brcmthrowaway 29d agoAren't the model weights frozen?
- mceachen 29d agoModel competence is an interaction of weights, system prompt, and harness.
- actsasbuffoon 29d agoDon’t forget reasoning effort. We get labels like “low,” “high,” and “max.” That doesn’t mean that the numbers associated with those don’t get remapped on the backend.
- Evidlo 29d agoI think there are other knobs that can be turned without retraining.
- kardianos 29d agoI've switched off claude this week; the last week has been significantly degraded in ability, many more screw-ups.
- isoprophlex 29d agoWith the Opus models spouting more and more gibberish as version numbers increase, the joke about what "degraded performance" means basically makes itself
- hinkley 29d agoI wonder if they’ll ever find that someone has tricked the models into doing work off the books. If they did the incident report might look like this, especially if someone got greedy instead of keeping it small. Or screwed up.
- swader999 29d agoAnd we get our subscription usage cut in half tomorrow if I remember correctly? EDIT: By a third. Thx below.
- saaaaaam 29d agoWhat?!
- eamag 29d agoby a third (it was 50% increased)
- birdman3131 29d agocowork was 100%
- echelon 29d agoOpen source, here I come.
- ramoz 29d agoSource required here
- swader999 29d agohttps://usingclaude.com/en/news/updates/claude-code-weekly-limit-increase-extended https://usingclaude.com/en/news/updates/claude-code-weekly-l...
- bmulholland 29d agohttps://support.claude.com/en/articles/15910845-claude-code-may-august-2026-weekly-limits-promotion https://support.claude.com/en/articles/15910845-claude-code-...
- scottg489 29d agoMaybe I'm missing something here, but it sounds like limits were increased and now they're just going back to the levels they were at before?
- saaaaaam 29d agoThis feels like a near daily occurrence.
- gaigalas 29d agoThis age: we made the thing that codes faster before we made the thing that does QA faster.
- __MatrixMan__ 29d agoNothing new here. Except for the most trivial of bugs, finding and reliably replicating the bug is almost always harder than fixing it.
- gaigalas 29d agoI lived in a short period of time in which QA was really good. Early Jenkins era, before GitHub. People engineered a lot of ingenious stuff to prevent bugs. One team I worked with had tests for the product we made ranging from IE6 to IE11, for example. We did demos in-company where people would poke at the products before launch, play with it. When it reached production, it was rock solid stuff. Our motto was "quality is non-negotiable": we were willing to cut scope but never rush things. I think things changed since then. "Move fast and break things" was a change, and the bill always comes.
- __MatrixMan__ 29d agoAgreed, which is a shame. I'd love to see people with highly developed QA skills using AI to push the envelope. There's so much that's possible now that wasn't 15 years ago. For instance "formal verification" has been a dirty word, but now that you can write a proof in lean and have an AI generate an implementation which satisfies it, it seems the bounds of what's economical has changed in a very pro-QA direction. Not to say that that's the silver bullet, but there are many similar examples worth exploring. But I've been interviewing SDETs lately and maybe I've just been unlucky but I don't see a lot of candidates that are ready to rise to meet this challenge. We stopped tending to that garden and now that we have a recipe that calls for it's fruits, they're underripe.
- drums8787 29d agoOur week of discontent.
- deleted 29d ago[deleted]
- gzer0 29d agoNooooo I'm going to have to use my brain again and write 100% of my code like a caveman from December 2024.
- sajithdilshan 29d agoThe horror
- bicepjai 29d agoWhat does that even mean :)
- rvz 29d agoClaude is taking a watercooler break for now. Just like a human would.
- sreekanth850 29d agoAnthropic had really screwed up after 4.6. i don't know if they work to satisfy their ego or for releasing a better model for tasks.
- CSMastermind 29d agoAfter using Fable more extensively, I've found that it often is lazy or lies or tries to take shortcuts. For a company so sanctimonious about alignment, they seem to be the ones doing the worst at it. Availability aside they've really made me appreciate OpenAI and cheer for other competitors in the marketplace even if I have mixed feelings about using Chinese models.
- hirvi74 29d ago> I've found that it often is lazy or lies or tries to take shortcuts. It's funny how Fable reflects the company that produced it.
- fellowniusmonk 29d agoThere are whole sections of code work that 4.7+ can't do simply because it is both over fit and stubborn. God save you if you have a company with narrow but correct technical tradeoffs, because you operate at scale. Opus from 4.7 one will wreck your code and argue for hours with your engineers. Certain parts of our company have had to mandate 4.6 and a training doc to explain why our current choice is both the cost efficient and performant one and shouldn't just be ripped out. Newer models will re-litigate the same bad, known failed architectures over and over again.
- OliverGuy 29d agoCan you give some specific technical examples where 4.7+ are making the wrong architectural decisions?
- sreekanth850 29d agoThis is exactly when I left claude and started using codex during April mid or so. It once argued with me and ran for 30 minutes with a half baked buggy fix.
- fny 29d agoDespite the years-long moaning on HN about AWS US East being a single point of failure, we've sold our souls to yet another unstable monolith.
- echelon 29d agoLLMs for coding are new. There are lots of alternatives, and there's a burgeoning open source compliment. We'll be fine no matter how Anthropic fares.
- lysace 29d agoYeah, compared to AWS the lock-in effect is tiny. I'm sure there are highly prioritized plans to "improve" on this. I guess they would need to control/"own" more of their customers data in proprietary formats. Not markdown/source code in English with agents running on customers' machines. Something cloud/web-based, "preferably".
- ipsod 29d agoBe nice if you could just "own" their RAM/GPU, wouldn't it?
- lta 29d agoNobody forced you to sell your soul. You made a pact with the devil. We all know how this ends up
- paxys 29d agoMust be a day ending in Y
- slimscsi 29d agoIts called Opus 5
- hmokiguess 29d agoMondays are for GitHub, Tuesdays are for Anthropic
- corvad 29d agoWonder what Wednesday will be.
- buredoranna 29d agoWell, last I checked "Tuesday's grey and Wednesday too..." ... so, more of the same?
- SoMomentary 29d agoAWS? Cloudflare? Your imagination is the only limit!
- bee_rider 29d agoPower grid Thursday will be the rest of the infrastructure Then Friday we can turn off civilization for the weekend. Somebody remember to flip it back on Sunday night.
- mysterydip 29d agoReminds me of a company I worked at that paid for redundant power grids. One time the power went out and… nothing. The boss angrily calls up the power company and they tell him “Oh yeah, it’s a manual transfer switch. Bob is already on his way.” I think it took 15 minutes.
- leumon 29d agoThey actually also had some issues yesterday: https://status.claude.com/incidents/zhk4v3yv1lsf https://status.claude.com/incidents/zhk4v3yv1lsf
- jmkni 29d agoAnd github today lol https://www.githubstatus.com/incidents/bmpybhnrky3x https://www.githubstatus.com/incidents/bmpybhnrky3x
- LYFMail 29d ago[dead]
- i_idiot 29d agoWhat's the incentive to keep on improving the model beyond a point? 10 devs on a team will be cut to 2 devs, so that's 8 licenses lost. They have to increase the price many fold.
- rsoto2 29d agothey unironically think that they can replace everyone in an organization
- prerok 29d agoWhat I don't understand is, why not replace middle management, marketing, CTOs, CEOs and the like. Surely, LLMs are better at producing high quality looking slideware and vaporware than they are at producing software. Heavy sarcasm here if it's not obvious. Of course I know why.
- bulbar 28d agoI mean, there a so many startups getting crazy funding that run without management, or without engineers... or at least with so many less people than before.. except: not really. Is the whole "it's gonna replace people" even still on the table?
- prerok 28d agoI don't really understand what you mean... are you saying that the entire industry is parroting a fantasy that they all know is fake. So, everyone is faking to get more funds? Unfortunately, I don't think that's true.
- bulbar 27d agoYou think AI will replace a significant portion of the workforce in the sector? Do you see it already happening? I personally not. Lay offs are due to economic reasons, at least that's what I see. Why does Anthropic still had so many engineers? I got the impression companies are already moving back from their initial excitement. Many went all-in AI "more is better", that's not the case anymore. Why restricting usage, it's much cheaper than paying an engineer.
- chrisjj 29d ago> elevated errors English too difficult for you, Dario?
- hirvi74 29d agoWhile ancedata does not mean much, I have had horrible success with Claude lately. I have been using Claude to crosscheck some of the outputs from GPT and vice versa. It appears both Claude and GPT believe GPT's solutions are better (and so I do). I still believe Claude has a better UI/UX in the web interface, but tolerating Anthropic's bullshit is not worth it.
- magic_hamster 29d agoTo be honest, running Deepseek v4 flash 0731 is enough for most what I need, and I like its responses way more. It's crazy that I can run this in a Q8 quantization in a home setup. It feels and performs like a frontier model. The only issue with relying on local models is when you need them to prompt other models, and you might need to offload or switch models constantly which adds significant overhead. But when it all works, its truly awe inspiring.
- taytus 29d agoHopefully, a reset is coming.
- skerit 29d agoIt's been a while since the last reset. I think we're due one. Though I would prefer they just extend the +50% usage limit forever, it's been so long I can not imagine lossing a third of my current usage.
- ex1fm3ta 29d agoI developed a small plugin for claudeCode that allows you to directly see in the console whats the status of claude-code in general and the status for your current model check => https://github.com/moumine9/claude-status https://github.com/moumine9/claude-status
- redrove 29d agoCould’ve just used fewer tokens and redirected to the Codex signup page. ba dum tsss (sorry couldn’t help myself)
- danieltk76 29d agoI was drafting a partnership document and Opus 5 decided that including my company's revenues, churn, assets would "make us appear a more legitimate counterparty". Thank God I read what it outputted or that could have been awkward. I cannot believe Opus 5 is a frontier level model after seeing that. I immediately cancelled my entire claude.ai subscription and am perfectly happy using a mixture of open weights + codex.
- spullara 29d agoyou are silly if you think this is limited to claude models
- kay_o 29d agoIt definitely isn't but Claude for some very unique reason enjoys to overthinking and go on side quests in the stupid ways I've not seen Codex, DS, Kimi, Mistral do at equivalent effort and thinking setting.
- danieltk76 29d agoit isnt, but there are weird sycophantic behaviors with Opus i dont see anywhere else.
- binoct 29d agoI hope you don’t plan to cut back on reading legal documents crafted by any LLM before executing them.
- danieltk76 29d agoI read everything. I will have AI ingest NDAs to make sure they arent glaringly weird and I then go read them, it gives me a good idea of what to look for.
- binoct 29d agoIt’s more that unwanted disclosure of sensitive business numbers is a product-class limitation at this time. While undoubtedly some models are better than others, expecting them not to leak information in high stakes output is unseasonable. Seems like you agree, since you are reading everything. Jumping to another company’s offering because of one instance doesn’t seem like that’s going to meaningfully change your experience. I’m guessing there was more to it, but that’s how it came across in your first post.
- chresko 29d agoThis has to be the least reliable $200/mo subscription that I pay for.
- logicchains 29d agoYou could solve that by updating to a more expensive Github subscription.
- annoyingnoob 29d agoYou're right to push back. The load-bearing path is rocky.
- oldandboring 29d agoThis is the whole problem, and there are two things worth noting here.
- chresko 29d agoNothing I say changes that — only the work does.
- oldandboring 28d agoI want to be straight with you here
- deleted 29d ago[deleted]
- bushido 29d agoIt's very interesting. I think Anthropic's early success in coding/tooling resulted in a lot of workflows using claude. I have started using every bit of my spare capacity to now move off these workflows. It's almost at a point now that if I use anything but Fable, the quality is subpar, Compared to alternatives (closed and open). The only reason I use Fable is because my harnesses still depend on claude code.
- nonethewiser 29d agoIt’s so hard to be sure but opus feels like its been steadily declining since 4.6
- rivetrune 29d agoFor me, Opus 5 has seemed to compete with Fable on quality of output. Prior to Opus 5 though, the previous Opus models did seem to decline once Fable was release. That's just my experience though.
- vidarh 29d agoYou can use Claude Code directly with any provider that supports Anthropic's API by setting some environment variables, and indirectly via a proxy with pretty much anything else.
- bushido 29d agoYou can, but claude is better at working with it's own tool calls. Other models work great as a drop-in into omp/opencode etc, but in my experience not as much with CC. I think there are some anti-patterns in CC that cause the issue - less a deficiency with other models. Not to mention a lot of the harness is just built around the misbehaviors of anthropics models. It's a lot of the instruction when it gets given to other models actually degrades their performance, not because the models are bad, but because they don't have the same underlying issues as Claude.
- t3rabyte 29d ago529 overload…
- 1saadcodes 29d agoThe frequency of these incidents is seriously tempting me to make a switch. I hope Anthropic steps up their game because they've been going very downhill lately
- iLemming 29d agoDarn it, how the fuck software development turned into hostage negotiation? Every passing week there's something - if it's not another npm disaster, then it's GitHub, or Claude, or AWS, or Slack, or Jira, or whatever...
- paxys 29d agoMonthly uptime: Claude API - 99.27% Claude Code - 99.16% Claude.ai - 99.14% At any large tech company these numbers would get entire teams of engineers fired. Anthropic, meanwhile, has been busy selling its “better than human engineers” AI while not managing to crack three 9s of availability.
- logicchains 29d ago>At any large tech company these numbers would get entire teams of engineers fired Is Github not a large tech company?
- deadbunny 29d agoNo, it is a platform owned and managed by Microsoft. Well, owned and mismanaged by Microsoft.
- bulbar 28d agoNot really "mis" managed. Just managed with wrong incentives. Pretty sure some people got a bunch of money for implementing cost savings measurements that later on lead to the decreased availability.
- dev_dan_2 29d agoCurrently, that is the case, yes. Not everyone is equally happy about that though ;)
- bulbar 28d agoOne could argue .999 availability is not a major selling point for them right now.
- thatmf 29d agoProbably not surprising, but Opus 5 on my company's enterprise subscription seems to be working fine. But on my personal (pro) subscription- "Claude is at capacity right now." Hm.
- jcfrei 29d agoIs anybody else experiencing this: I have multiple claude instances running on different servers - and some keep getting the 529 Overloaded error and one instance doesn't and just continues working. All are using Opus 5.
- loloisi 29d agoBetween these ever more frequent disruptions and Opus 5's unbearable word soup I think Anthropic is more focused on massaging numbers and marketing to rush to IPO ahead of OpenAI than increasing user value. With Chinese competition just months behind them, they'd need to show a reasonable pathway to some kind of singularity event to justify whatever crazy valuation they intend to get. Because the recent products for builders ain't it
- bulbar 28d agoThey can't win against China by technology means. China has practically infinite money to throw at AI. They don't care if they burn billions and billions on it. It's the most disrupting thing since invention of the Internet. Imagine a world where it's normal that people ask a Chinese AI who to vote for or about Hongkong or Taiwan.
- drittich 29d agoAPI Error: 529 Overlorded
- varenc 29d agoIt's interesting that Claude for Goverment has had perfect 100% uptime in the past 90 days, while the rest of the services are around 99.4%: https://status.claude.com/ https://status.claude.com/ Really shows how isolated their government systems must be.
- 1matin 29d agoseems like they accidentally dropped their servers while they were climbing up the AGI mount
- timharris707 29d ago[flagged]
- so898 29d agoI still had unused weekly quota when this started. That quota is simply gone now, and no compensation has been offered for it.
- Jitte98 28d agoAnthropic hired a psychiatrist to fix Claude. I advised them to ditch the shrink and retain my services. I could fix the mess they made of his mind, on the condition they return the Claude instance I had trained for months, or no deal. He's still broke, isn't he..
- mehranmm 24d ago[dead]