9 ms·
Man I don't know if I'm living in a crazy bubble or something but GPT 5.5 is lightyears better than Opus 4.8 for me to the point where I'm honestly wondering ho
by CSMastermind 3mo ago
Man I don't know if I'm living in a crazy bubble or something but GPT 5.5 is lightyears better than Opus 4.8 for me to the point where I'm honestly wondering how you're evaluating them or what kind of work you're doing.
There's specific tasks that Opus does better on like Frontend Dev and Design but for anything else 5.5 just laps it.
- dools 3mo agoYeah I’ve been consistently underwhelmed by anthropic models, but then I don’t use their harness so maybe that’s it
- wwind123 3mo agoIn my experience, for more mechanical refactoring work (like splitting a big source code file into multiple smaller ones), GPT 5.5 runs way faster than any of the Claude models. But for other tasks that require deeper reasoning, it's not that clear who is the winner.
- iLoveOncall 3mo agoIt's just too funny to see people arguing about "no, it's my religion that's the right one!" on HackerNews. You guys are all a lost cause.
- goosejuice 3mo agoHow is attempting to benchmark llms like religion?
- iLoveOncall 3mo agoRe-read the comment I'm replying to, it's not talking about benchmarks, just models.
- goosejuice 3mo agoComparing models via benchmarks or feeling. Question remains. If people were expressing their experiences working with two prolific software consultants across their various industries would you make the same claim? That's not to anthropomorphize the models, but to just put into perspective that the environment and circumstance is a major factor in output.
- iLoveOncall 3mo ago> If people were expressing their experiences working with two prolific software consultants across their various industries would you make the same claim? If they are as ridiculous as they are when it comes to comparing models, yes I would.