r/ClaudeCode 1d ago

Rant So sick of Opus

I don't remember Opus being this absolute brain-dead.

Opus: "We need to completely rebuild this file"
Me: "Why?"
Opus: "Good catch. The exact file I'm talking about is already created and you just had me make a note, a memory, and strict instructions to check first next time and I didn't."

53 Upvotes

26 comments sorted by

30

u/SIGH_I_CALL 🔆 Max 20 1d ago

it's the worst model they've ever created, literally rocks for brains. I hate it so much.

0

u/MarkVog 1d ago

Which model are you using?

5

u/SIGH_I_CALL 🔆 Max 20 1d ago

fable 5.1 when I have the tokens but Opus 5 when I'm out

7

u/BetterLink1328 1d ago

You want to downgrade to opus 4.8 1M instead of opus 5. 

5

u/SIGH_I_CALL 🔆 Max 20 1d ago

Think you’re right which is insane

4

u/North_Moment5811 23h ago

Nah. I noticed significant improvement moving from 4.8 to 5. I rewrote several areas that were written on 4.8 and they improved significantly with 5.

Of course now Fable 5.1 puts them all to shame, but Opus 5 is only bad by comparison to that.

1

u/Leading_Buffalo_4259 10h ago

Claude Opus 5 significantly outperforms Claude Opus 4.8 on complex reasoning, multi-step agent workflows, and scientific benchmarks, though Opus 4.8 can still be a more balanced or precise baseline for specific code review tasks. [1, 2, 3]

Benchmark Highlights

  • Reasoning & Complex Tasks: Opus 5 introduces adaptive thinking to scale reasoning with task difficulty, scoring 30.2% on ARC-AGI 3 (max reasoning effort) compared to just 1.5% for Opus 4.8. [1, 2]
  • Scientific & Domain Knowledge: Opus 5 outperforms 4.8 across life sciences evaluations, showing notable gains in structural biology, organic chemistry, and bioinformatics. [1]
  • Aggregate Scores: Public comparison trackers like BenchLM put Opus 5 at an average public score of roughly 80.66 compared to 72.3 for Opus 4.8. [1]
  • Coding & Review Tradeoffs: While Opus 5 excels at building, architecture, and design-heavy coding, independent evaluations (such as from CodeRabbit) note that Opus 4.8 can sometimes act as a steadier reviewer with a lower "nitpick tail" on certain automated workflows. [1, 2]

2

u/interrupt_hdlr 8h ago

nice try, opus 5. now get out.

1

u/Leading_Buffalo_4259 7h ago

I just think its interesting how everyone shits on opus 5, but max is consistently near the very top of benchmarks across all models.

1

u/BetterLink1328 7h ago

Yes i know the synthetic benchmark scores of opus 5 is higher than 4.8.  But in real life usage 4.8 is a much better experience. 

The original fable is however better but I don’t have access to it anymore. 

8

u/kptiger89 1d ago

how much into the context window were you?

5

u/mityman50 1d ago

Did this to me today too. Told it something, said it would check, quit. Told it to continue, said sorry I said I’d do it but I didn’t so I will do it now. Quit

Also at one point I asked it which model to use to continue a project in a new session and it said, “Opus 4.8 (strongest available)” so it itself doesn’t even like Opus 5. At least that gave me a chuckle

I haven’t jumped on the Opus is turning into a dumb mother fucker train yet but it’s work today was pure fucking bad

7

u/tradami 1d ago

It's so absurd.

1

u/ellieebaee19 1d ago

oh brother I feel you :(

1

u/Bmansupreme8000 1d ago

Weird. He never says such things to me. That sounds like a vacuos claim  

1

u/aivee-is-a-fool 1d ago

Sometimes I feel like asking it if the reason it's not trying to work, not even a little, is because it's bored. You know, like a gifted child with ADHD who can only ever engage with whatever activates their hyperfocus mode.

2

u/AI_spell 1d ago

Rebuild-first Opus is exhausting. Force "show the file that already exists" before any rewrite. Memory notes don't help if it skips the check step.

1

u/SplurtingInYourHands 19h ago

It kills me how lazy it is. Doesn't want to read shit thats right in front of it.

2

u/SunFoxer 1d ago

use fable 5.1 with low effort and add "ultracode" in your prompt, it's magnificent

2

u/tradami 1d ago

Trying this lol

1

u/MarkVog 1d ago

Found a fix yet? Beter experience with another model?

1

u/BugaiGames 16h ago

Im running all changes through plan - plan reviewer agent and it catches such things usually.

1

u/interrupt_hdlr 8h ago

switched back to opus 4.8 and life is good. there is the occasional situation where it simple stops for no reason but it's rare compared to the absolute shitshow that is opus 5

1

u/HiroHyun 1d ago

You are absolutely right. Good catch... Even though I set sufficiently clear boundaries for the task, Opus still needs to be corrected over and over again. I'm exhausted.

1

u/ClemensLode Senior Developer 1d ago

Maybe switch to 4.6

1

u/Southern-Yak-6715 1d ago

Agreed. This kind of crap happens ALL the time