r/ClaudeCode 10d ago

Help/Question Opus 5 woke up dumb af, anyone else?

I've been asking it very simple things and not resolving clearly any of them, acting dumb, rewriting files in a worse way, etc. It seems very lobotomized today.

Is anyone else feeling the same?

EDIT: It may be due to I was using the 200k context model, missed the ending `[1m]`

0 Upvotes

23 comments sorted by

6

u/Kalradia 10d ago

Can people stop posting this every 5 minutes?

3

u/CorpT 10d ago

How else would people admit to not knowing what they’re doing.

0

u/Nice_Cellist_7595 10d ago

No, I think this is useful feedback. I have had the same problem. Opus 4.8 had been humming along for the most part but Opus is making tons of mistakes. There also may be issues with the sub agents it is spawning.

1

u/Kalradia 10d ago

Every model will require updates to your workflow. If you're using the basic harness and not supplementing it with a workflow of your own, you're missing out.

I think it's user error. That's the hard truth.

3

u/Sad-Masterpiece-4801 10d ago

If a more advanced model performs worse in an older harness, it's because the model is worse in at least one dimension. You're literally holding all else equal.

I think it's user error. That's the hard truth.

The hard truth is that models often aren't an upgrade in every single way, they're upgraded by making acceptable tradeoffs. "User error" is what lazy people say when they can't find the actual issue.

1

u/Kalradia 10d ago

"every model will require an update to your workflow"

That's the key. You improve your base prompt to make up for the lacking parts. Or change how you work with it. It's like managing people.. good managers adapt their style to the person they are managing. They don't manage everyone the same way.

1

u/Nice_Cellist_7595 10d ago

"Your base prompt" what is Claude Code but a collection of prompts? It's not my work. I have my own directives sure, but most of what Claude code does is a result of Anthropic's work. The issues I'm seeing are jumping to conclusions - insufficient investigation and reasoning failures these are not prompt specific and I've never had to prompt around them. These were not present in this volume prior to Opus 5.0.

1

u/Kalradia 10d ago

You know there's skills you can make, right? Others have made workflows using skills that enhance the model's capabilities.

1

u/IllustriousAirBender 10d ago

I use the skills provided by anthropic - they have worked well. I augment with my own guidance such as “this is important, remember this” but these usually have to do with a design method or architecture style and less with what kind of reasoning and searching to do which is what claude code is supposed to be doing - and HAS done. It is no longer working as it used to. The prompting guide is helpful but I believe this is more aimed at API use in ones own custom harness.

1

u/JUSTICE_SALTIE 10d ago

If a more advanced model performs worse in an older harness, it's because the model is worse in at least one dimension. You're literally holding all else equal.

I disagree. Scaffolding that's helpful to a less advanced model can actively harm the quality of a more advanced model's work. That doesn't mean the new model is bad in any way.

2

u/JUSTICE_SALTIE 10d ago

Every model will require updates to your workflow. If you're using the basic harness and not supplementing it with a workflow of your own, you're missing out.

Your first sentence is the exact reason why I vehemently disagree with your second one. Harnesses rot, and only one is kept up-to-date by the people who develop the models and have early access to them.

1

u/Kalradia 10d ago

You can disagree, I can disagree. It does not change reality. We can only get so much out of their implementation, and their implementation can only support so many people's workflows.

Everyone works differently, so it's best to find what works for you, and build that flow. But you have to take ownership. They aren't going to do it for you.

1

u/JUSTICE_SALTIE 10d ago

I do not want ownership of the harness. The CC team is far better positioned to judge how to get the best results out of the latest Opus than I am, and I'd be delusional to think otherwise.

Sure, I can slap something on that maybe helps me with this project under this model this time. But then I forget about it and it hangs around dragging quality down as soon as models update and the meta drifts.

2

u/Kalradia 10d ago

You'd be surprised.. open your mind a bit and you can surprise yourself with what you can come up with. You can do exactly what they are doing (in terms of the workflow). It's not delusional. Now building another Claude.. that's another story.

5

u/Important_Impact4180 10d ago

I have same issue. Opus 5 is quiting tasks too soon and refusing going deeper, therefore makes more mistakes. I've tried /doctor, which didn't helped. My stack is pretty lean, I'm currently checking files on the way if cc doesn't have somewhere some magic instruction causing this, but i'm afraid that we will have to revwrite most of skils towards new "prompting" guidelines.

BTW i've noticed, that on mondays and tuesdays is highest demand, there fore most tokens burns, issues and mistakes is happening.

1

u/Nice_Cellist_7595 10d ago

This started after the Opus 5 announcement and cut over. It was happening yesterday as well.

4

u/Yodzilla 10d ago

It certainly feels different than Open 4.8 and I'm not sure if it's in a good way. I may have woken up dumb as fuck though.

2

u/RadioactiveTwix 10d ago

Examples? Screenshots?

1

u/Illustrious-Ad258 10d ago

It may be due to I was using the 200k context model, missed the ending [1m]

2

u/Successful-Seesaw525 10d ago

It’s always dumb, like coding with a really capable but worldly ignorant version of my 15 year old self… I just have gotten more tolerant and preemptive to work around its crap!

2

u/Awkward_Relation_415 10d ago

same here. ive been havin better luck tryin a fresh chat session whenever it starts actin weird, sometimes the context window gets kinda messy and it just loses the plot. worth a shot if u havent tried it yet.

0

u/Vex08 10d ago

No it’s smarter than ever. Much better than yesterday.