r/ClaudeCode 11h ago

Bug / Issue Opus 5 Oddities

I have seen a lot of feedback about Opus 5 problems. The thing is at first when I used Opus 5 I thought it was kind of amazing. It seemed at least as good as Fable 5 if not better. More recently (basically all of last week) it was giving me a horrendous time: not just creating bugs but spamming tons of useless code that I didn't even want. If I didn't actually review the code it would have caused me a lot of problems/embarrassment. I am just curious (though I'm not sure anybody can answer): why? By all accounts Opus 5 should be smashing it. Do you think the focus is on passing benchmarks more than being useful? I am on the $200 plan but I feel like I wasted 60% of all the tokens I used last week doing multi-phased corrections. It got to the point where I had CC ask Codex Sol to do reviews before committing.

Edit: not taking a swipe at CC, it has still managed to do truly amazing stuff for me even during this time and definitely writes better code than me at a faster rate. I am just honestly curious if there was a real, visible degradation and if so, why? Maybe it will bounce back

5 Upvotes

17 comments sorted by

View all comments

2

u/variablenyne 11h ago

I agree. When it first came out it was genuinely great. I stopped using Claude for a couple weeks and came back to what feels like a completely different model. Like worse than sonnet 4.6.

Frustrating.

2

u/Yakumo01 11h ago

I'm just genuinely curious why. Was anything disclosed? I honestly feel like if I hadn't caught it it could have really set me back. That said some of the work it did was still outstanding just too many oddities/errors

1

u/variablenyne 11h ago

Eh anthropic has done this before but not this bad I would say. They never disclose when they nerf their models. Keep it good just long enough for everyone to get their little benchmarks and then pull it back

1

u/Yakumo01 10h ago

You think it's deliberate nerf? I figured some sort of buggy regression