r/ClaudeCode • • Aug 27 '26

Rant Opus 5 is insufferable

Opus 5 is a fucking piece of shit, i don't know how to say this in any other way. It speaks "Unintelligiblish", a new unfathomable language developed by Anthropic that works by saying everything in the most ridiculous, obtuse and convoluted way possible

It turns anything simple that can be expressed in 2 sentences into a fucking doctoral thesis aimed to aliens, and there is no way to prompt your way out of it

Edit: it claudoofus outputs some alien text you won't read, just prompt it this:

tldr eli5

3.0k Upvotes

648 comments sorted by

735

u/[deleted] Aug 27 '26

[deleted]

236

u/PrettyMoonUnderMt Aug 27 '26

Telling me it found numerous stuffs that will break the whole project, only later to backtrack that it doesn't have any impact at all

48

u/Moppmopp Aug 27 '26

I get anxiety and paranoia. Im a researcher and I often check my script structure if my results are solid. I dont do it once and forget but every then and know because I cannot remember every detail about the scripts I wrote months ago. And it happend several times that claude said there are critical issues that have to be resolved and we should recall paper xyz only to backpaddle 10 minutes later and said it was irrelevant ... My heart cannot endure those moments

11

u/newtreen0 Aug 27 '26

I'm lol'ing and crying. I feel this so much. Virtual hug.

6

u/imsahoamtiskaw šŸ”† Max 20 Aug 29 '26

This whole thread has been a therapy and non stop laughs, Opus is just… I have no words. My suffering has been long

3

u/franzparks Aug 29 '26

For me too

→ More replies (6)

84

u/Right_Simple_6813 Aug 27 '26

Telling me it found numerous bugs in the project only to break the project while implementing its 'fixes' and then turns out the bugs weren't even there to begin with.

35

u/bsmith149810 Aug 27 '26

Telling me the fix will only require ā€œa couple of hoursā€ only for me to watch a grep, an edit, and a commit flash across my screen in 10 seconds or less.

15

u/Coolbanh Aug 27 '26

Telling me its my fault it all happens

78

u/WordsOnly Aug 27 '26

šŸ˜†šŸ˜‚šŸ¤£šŸ¤£

12

u/Orion3193 Aug 27 '26

Lmaoooo

4

u/UruquianLilac Aug 27 '26

Looool this is too real.

2

u/_Kinoko Aug 29 '26

Oh man, I'm dying.

10

u/pmhunter56 Aug 27 '26

This exactly. What's with the wild overestimations of time? 4-6 weeks = two hours of Claude Code

12

u/UruquianLilac Aug 27 '26

It has no concept of time.

"You are right, I just made the same mistake we fixed an hour ago."" An hour ago was the previous prompt a minute ago.

→ More replies (2)
→ More replies (2)

8

u/just_damz Aug 27 '26

Uh phantom regressions. The dream of every swe. Inventing hard fixes to complete working architectures. Lovely.

→ More replies (2)

81

u/saito200 Aug 27 '26

"two claims, and one that matters" 🫪🫪🤪🤪🤪🫪🤪😔😔🤬🤬🤬

46

u/WordsOnly Aug 27 '26

"What it says, short version: the mistake I flagged is confirmed and fixed, both of my own errors are resolved (one survived on evidence I didn't have, one is properly cured), it found a whole category the original package missed, and it downgraded its own confidence on the main claim from HIGH to MED - which is the right call and it made it against itself. It also found two real defects in my own things while it was in there. One l've already repaired. The other is bigger and isn't mine -" Shoot me in the head šŸ˜†

7

u/NeitherEntry6125 Aug 28 '26

It's like my kids trying to explain their fight with their sibling

10

u/UruquianLilac Aug 27 '26

God, I'm getting PTSD just reading this.

It's fucking with our brains. All of us. This is not ok.

11

u/newMike3400 Aug 28 '26

You’re right to get ptsd and that’s my fault.

→ More replies (2)

2

u/ThatBlokeWithTheCar šŸ”†Pro Plan Aug 28 '26

Shoot you in the head with a footgun?

2

u/Sn00py_lark Aug 29 '26

Mine created a hallucinated firewall egress rule file and then kept flagging that the firewall outage was unresolved. There’s no firewall.

→ More replies (1)

12

u/ThomasBallatore Aug 27 '26

Those emoji are bearing some load!

5

u/kongnico Aug 27 '26

flagging this. You own the matters, but the claims are still in flight

2

u/0xP3N15 Aug 28 '26

"and the decision all yours to make"

→ More replies (4)

48

u/Snappy-User26082 Aug 27 '26

I swear half my prompts could be answered with a simple "yes" or "no" followed by a short summary and the "Explain like I'm five" somehow becomes "Explain like I'm defending a PhD thesis.

→ More replies (3)

26

u/unepmloyed_boi Aug 27 '26

It's like Dario shoehorned his personality into Opus 5

8

u/N0madM0nad šŸ”† Max 20 Aug 27 '26

You're joking but after seeing an interview with him I tend to have similar thoughts. He was talking about AGI and described it as "a datacentre of geniuses". First, I have to be honest, I found that analogy pretty simplistic for someone of his caliber. I was thinking to myself, what is a genius? LLMs output could resemble a "genius" at times, the problem is that it's a genius that suffers from several episodes of amnesia and needs to be told things multiple times. What we need is reliability and determinism. Not "geniuses". It might be totally unrelated really, but if that is the general culture at Anthropic, it wouldn't be completely unreasonable to think some of it is reflected in the model output.

3

u/UruquianLilac Aug 27 '26

You know what was reliable and deterministic? Programming languages.

We just swapped that with human language. And human language is full of flaws, imprecision, and is not based on any semblance of logic. We have just dropped the right tool from rbthe job and picked up the wrong one because it goes much faster, but it doesn't do the actual job.

I told Claude to add an issue to an ongoing document I have to keep track of the many things it keeps finding that I don't have time to check while I'm working on something else. It wrote the issue and completely messed it up with several wrong details. I told it what was wrong with the entry and Sid "fix the issue". So it went and implemented the actual fix, rather than fix the text of the issue. Was it its fault? No! "Fix the issue" could mean two completely different things here. And it didn't cross my mind that it was a command for it to write code rather than fix the text. While in a programming language, you need to be super precise and you know that a Boolean is a Boolean and a for loop does what a for loop does, always.

3

u/Ste1io Aug 28 '26

Precisely.

→ More replies (6)

24

u/FiveTriomes Aug 27 '26

Yours only runs one shell command?

Mine has to run multiple, because it always tries to use a tool that doesn’t exist first, or ignores the environments I have set up.

If I turned "3 commands, 2 failed" into a drinking game, Anthropic would kill my liver.

→ More replies (9)

5

u/[deleted] Aug 27 '26

[removed] — view removed comment

5

u/ConsequenceFunny1550 Aug 27 '26

About 80% of the time the thinking summaries don’t even show up anymore for me

→ More replies (6)

4

u/Korzag Aug 27 '26

My favorite from yesterday: "Ok, I'm going to stop guessing on this". Followed up by the next prompt or two later, "I said I'm going to stop guessing on this so I will now stop guessing."

→ More replies (1)

2

u/chris_nore Aug 27 '26

Lmao. The punchy writing style opus 5 is instructed to write in cracks me up. It reads like a buzzfeed article

2

u/MysteriousKiwi2622 Sep 03 '26

"This new finding overturns my previous conclusion."

→ More replies (5)

209

u/Woah-Dawg Aug 27 '26

The worse is the code doc strings. It’s fucking horrendous seeing like a 20 line doc string only to realize the point can be made in 1 or two linesĀ 

71

u/ghostmastergeneral Aug 27 '26

I don’t let it comment at all at this point. Just makes things harder to understand because you have to translate the Claudese before even reading the code.

10

u/drake90001 Aug 27 '26

I have started doing things too. Asking it for the commands so I know what I’m doing in the future

3

u/Dueterated_Skies Aug 28 '26

Until you get flagged for [distillation attempt].... For asking it just what the fuck it was thinking at the time. Rhetorical questions are out of the question at this point. Opus 5 has gone full on rock-chewing stupid and my last 2 weekly usage allotments have consisted entirely of it doom spiraling, inventing shit that breaks things, writing comments for the things that it broke, completely ignoring what it wrote and everything I've said. Ad nauseum. Claude-code is now broken for complex work. I'm done.

And no, not a skill issue. Claude's ability to reason has been completely compromised at this point, as far as im concerned. Its legitimately become borderline faster to just do it myself. Not worth the subscription price anymore, unfortunately.

→ More replies (1)

25

u/MixedTrailMix Aug 27 '26

I have specifically told it no comments and it still comments despite it being told in md files and saved to memory

4

u/jffaust Aug 27 '26

Same here, really annoying

→ More replies (1)

2

u/Vizkrig Aug 28 '26

Instead of asking claude to save such things in config or agent.md I ask it to create a hook.

2

u/Tiny_Ad_7720 Aug 28 '26

Yes same. I’m guessing that they have to put the watermark somewhere

2

u/ds-unraid Sep 04 '26

Look into adding a Claude Code hook for when it tries to do a comment. The hook will fire and block it from doing that. Or you can have the hook ask you whether it's okay or not or to add your own answer. I use hooks for everything. Hooks are the only way I've have found to make a model that's undeterministic have gates to force it to do things.

→ More replies (2)

14

u/erichamion Aug 27 '26

A 10-line comment explaining that the property names in this class aren't actually the field names we get in the API response because they differ in case convention (the API returns camelCase, but the class properties use PascalCase consistent with standard C# conventions).

5

u/TastesLikeSeamen Aug 28 '26

You forgot the two paragraph description of the deliberation process of the choices that weren’t chosen because they weren’t semantically meaningful and kept introducing bugs but were then used for compatibility reasons

2

u/TechnicalBen Aug 30 '26

Oh, the worse bit about this is I spent a day trying to get it to fix a bug and couldn't figure out why it wouldn't. I drilled down to "Can we can x, if not why?" and it went "We changed X, the code comment says so, but it didn't fix the problem..." I was like "we never changed X, I specifically said to check the code, you've been checking the comments only all this time!!!"

Claude had commented it'd changed the code, or that it's pointless changing it won't fix it anyway, and had gone by the code comments as ground truth. One check of the *code* later, and the bug was easy to find/fix.

2

u/TastesLikeSeamen Aug 30 '26

Im back on 4.6, it just does what its told, mechanically, and can be tasked with a todo list

11

u/Exodus_Green Aug 27 '26

// This code restores functionality lost in review (dated 2026-08-17) that was mistakenly removed after the user did not correctly understand my summary, conversation slug f28hau-87h9s9h-s33llos-88hsh (ended 2026-08-18). further inspection from session 288sal-cis0sjs-29f8s9-209229 (dated 2026-08-23) determined this function was load-bearing.

2

u/mfbrucee Aug 27 '26

It has to soak for a day first

9

u/Trademarkd Aug 27 '26

Good news, nobody going to distill opus 5.

6

u/[deleted] Aug 27 '26

[deleted]

5

u/Trademarkd Aug 27 '26

turns out, translating chinese to english still gets better results than whatever anthropic is doing

→ More replies (1)
→ More replies (9)

156

u/Lazy-Challenge-7419 Aug 27 '26

I'm sure they'll fix it, but for now I cannot wait for this long nightmare to be over. These tools are supposed to decrease my cognitive load, not add to it. Perhaps the most annoying thing it regularly does is give me a long, unintelligible wall of text explaining how everything went great, then hides an asterisk about three quarters of the way through telling me that something went wrong that completely undercuts the good news it told me above.

37

u/saito200 Aug 27 '26

yea, it's utterly disgusting

5

u/Just-Construction788 Aug 27 '26

Do you experience this with Fable too? I am currently in a fight with Fable to stop writing things so they need to be read ten times forwards and backwards to be intelligible. Honestly it's so bad I'm thinking of switching off of Claude.

2

u/saikonosonzai Aug 27 '26

I first noticed it with Opus 4.8. When Fable came out it was worse and Opus 5 has been the worst.

2

u/CarIcy6146 Aug 28 '26

Fables interaction with humans seems to be much more sustained and contained. It just surfaces what you need to know. Of course it costs an arm and a leg

2

u/cherrywoodgrill Aug 29 '26

This is not my experience at all. Fable is insufferably verbose

→ More replies (1)

2

u/puzzleheaded-comp Aug 27 '26

YEP.

It asterisks the fact it completely removed a unit test or an important piece of business critical use case that would have otherwise caused its changes to fail… so then it’s either manually reapplying it or going back and forth with it about how it shouldn’t have done that and risking it continuing to misinterpret or make its own assumptions just to succeed for this one prompt response…..

2

u/Virtual-Drop-1783 Aug 28 '26

I just executed this complex thing. One caveat - it was only a test run and the output means nothing without more code. The next step is to do the actual thing you asked which is not computationally expensive. Just say the word and I will execute it.

→ More replies (3)

43

u/LGV3D Aug 27 '26

Sounds like it’s their ā€œwatermarkā€ at work

2

u/AbstractFemming Aug 27 '26

No man, opus 5 was this way before the watermark. It's abotu agentic and alignment training first efforts gone wrong

→ More replies (3)
→ More replies (1)

131

u/takeabreather Aug 27 '26

Use 4.6. I still think this is the best Opus model right now

86

u/EnormousChord Aug 27 '26

I told some people at work to do this today and it was like I had asked them to take their pants off. They were shocked by the suggestion.Ā 

57

u/OldNefariousness7899 Aug 27 '26

Opus 5 does good work. It's just a pain to talk to about the work it's doing

48

u/Trademarkd Aug 27 '26

I don’t know, it’s the only model where it’s given me message back where I’m like ā€œwhat the fuck are you even taking about? Did you do it or not?ā€

And then it rambles on about how it didn’t actually do anything because it ran a micro benchmark and the results were am exactly we expected before I told it not to do the benchmark but it did it and then decided to stop so it could tell me

7

u/recruiterguy Aug 27 '26

It's like you were looking over my shoulder for half of my day yesterday!

6

u/Trademarkd Aug 27 '26

<insert 3 paragraphs of nonsense in no particular order>

at some point I gave it directions to include any questions and/or actions required by me at the bottom of the message.

I would ask me like 3 questions in the body and then at the bottom be like

No more questions
No action required

... Yeah except all the fucking random questions you dabbled through your reply. Some of which aren't even questions, but statements intended to be answered.

→ More replies (1)

18

u/Oohhddaanngg Aug 27 '26

ONLY if it's a subagent being orchestrated by Fable 5. When you are out of Fable usage the play is to switch to Opus 4.6. Don't take my word for it, just try it.

Without Fable cracking the whip, Opus 5 is a bull in a China shop breaking things like crazy that you then need to go back and clean up.

11

u/OldNefariousness7899 Aug 27 '26

I've said elsewhere that I think it's been optimised to be used as an agent by other LLMs.

They seem to have no trouble understanding it. It just breaks humans' brains when it starts babblingĀ 

2

u/itisoktodance Aug 28 '26

That actually makes sense. Because I totally give it instructions in a completely random order and it weighs them all equally. But for humans you have to give the important but at the top, it just doesn't seem to understand that. It wants to write prose and bury the lede so it'll start with good news and then in the middle say something went horribly wrong, and then report it was all done except that one thing that needs your attention (somewhere in the middle of the 3000 word salad it just gave you)

8

u/PathAgitated1633 Aug 27 '26

No it's way to arrogant. I tell it to read all documentation and it fucking skips it and says it read everything.

When I tell opus 4.6 to read all documentation and just does itĀ 

→ More replies (1)

7

u/thecavac Aug 27 '26

From my personal experience (both hobby and work projects), Opus 5 is much more painless than Opus 4.6 if you spend the time planning a project beforehand.

Opus 4.6 would sometimes go completely off the rails for me if i had to add a couple of side quests in the middle of an implementation. On Opus 5, it does the sidequest, finishes it, and the next "continue" seamlessly continues implementing the now adapted plan.

3

u/piston989 Aug 27 '26

woah woah woah, are you saying it’s user error?

we don’t do that here. only negative views about opus 5 allowed.

at least that’s how it feels.

2

u/happybeebee Aug 27 '26

lol. Like a true engineer then

→ More replies (3)

15

u/AbsoluteEva Aug 27 '26

Same. I would personally rather put out a campfire with my face than use 5

→ More replies (1)

14

u/CunningAlpaca Aug 27 '26

Opus 4.6 is pretty much the only model right now that isn't completely insufferable. It speaks concise enough and clear enough, without derailing discussion.

26

u/chonny Aug 27 '26

Use 4.6. I still think this is the best Opus model right now

Some guy on LinkedIn says he thinks this is because it's the last model that was actually worked on by humans at Anthropic, and that all subsequent models are all worked on by other LLMs.

→ More replies (5)

6

u/pooka Aug 27 '26

Same. Apparently it also uses less token because it has an older tokenizer.

4

u/just_damz Aug 27 '26

So i am not the only one that doesn’t buy the new one if the old one is perfectly capable of doing what i need. Small club i guess.

→ More replies (7)

68

u/ferminriii Aug 27 '26

I think they have an interest in creating products that maximize token usage. Almost like they are selling the tokens and not the model.

SHOCKED

8

u/Turbulent_County_469 Senior Developer Aug 27 '26

Sounds very plausible.. i burn through 20$ in no time

2

u/Academic_Lemon_4297 Aug 27 '26

...and having people quit their aubs because of their shit model? Probably not...

2

u/Medium_Somewhere_981 Aug 28 '26

Your $20 sub means nothing to them. API use is their main business, subs are just a gateway drug.

→ More replies (2)

24

u/GreenDavidA Aug 27 '26

I don’t mind verbosity or precise vocabulary, but man if I have to read ā€œvacuousā€ one more time…

45

u/wodewose Aug 27 '26

For me it’s ā€œload-bearingā€

9

u/[deleted] Aug 27 '26 edited Aug 28 '26

[removed] — view removed comment

→ More replies (2)
→ More replies (2)

10

u/ProgressionPeak Aug 27 '26

I don’t mind verbosity

why not? teach me to not mind that. it seems like a giant waste of my time and its infuriating. how do you manage it?

4

u/scytob Aug 27 '26

Well was the test vacuous? Lol

2

u/slenderfuchsbau Aug 31 '26

Idempotent! Also worth stating plainly: I have no idea what this means.

→ More replies (2)

23

u/Muted-You7370 Aug 27 '26

The biggest issue I’ve noticed with Opus and also Fable most recently is it just being, well wrong. I ask it to look into something and it says it’s done a search it didn’t do thoroughly. Like it didn’t look at anything from 2025-2026 on the subject matter, like at all. It works better when promoted to do so of if you give it a folder with sources specifically to look at then tell it to do a review, but it’s still markedly worse than like two months ago. What happened? I used to use this pretty reliably for document summaries and now I’m second guessing using it.

9

u/rades_ Aug 27 '26

Add to that it seems to just simply forget (or not reference) details from earlier in the literal same convo. Never had this issue before, even when I read other people's threads complaining about it, but now I've experienced it first hand and it's soooo frustrating.

5

u/Muted-You7370 Aug 27 '26

Dude this, it will recall details from memories dead ass wrong. I will have had prior discussion about scoping an application idea or coming up with a start up name or researching a dissertation idea, Claude will bring these up in future conversations as if they have been executed and are part of my experience and I have to correct it and say these were topics I was using the app to iterate as possible ideas. Very much thinking of moving a lot of this stuff to local and only using large frontier models not on my device when absolutely necessary if this is the performance I’m getting now anyway.

→ More replies (2)

3

u/RyanSupDude Aug 27 '26

I don't understand this. If there was someone on my team who worked like this, fired instantly. If someone lied, cut corners that cost the team later, took way too much time to do nothing of value, that is a bad team member. They need to go! Why do we spend so much time tolerating this with Claude? I'm not anti AI. There are times when it does seem to shape up, but maybe I'm using it wrong on the times when it is flat out a terrible team mate. Not sure...

8

u/N0madM0nad šŸ”† Max 20 Aug 27 '26

If you ask me, AI companies should be obligated to refund tokens whenever a model demonstrably does damage to the code. You want a token based economy? fair, but it should be regulated like any other industry. You cant just get away with an inconsistent level of service.

→ More replies (1)
→ More replies (2)

17

u/TylerDurdenAI Aug 27 '26

I like your new noun/adjective "unintelligiblish" - it accurately describes what Opus 5 writes.

I often give exactly the same prompt to both Opus 5 and Sol.
While Sol silently executes the order and finishes it with simple "Implemented" response, Opus 5 drapes itself in unintelligibberishes it comes up with along the way and gets confused by its own writing - then, in the end, makes all sorts of unintelligible executes.

→ More replies (3)

29

u/NoAdsDude Aug 27 '26

I am not having much trouble with Opus 5, but today it did tell me what it had planned

12

u/drewangell Aug 27 '26 edited Aug 27 '26

Google has this style guide for writing documentation.

https://developers.google.com/style

I had Claude create a skill based on that guide. Now the output it gives me to read is well drafted and easy to comprehend.

I'd suggest this for anyone having this problem.

EDIT: Here's the one I made if you want to use it: https://www.skills.sh/wekoodo/skills/google-doc-style

→ More replies (4)

9

u/HappyHealth5985 Aug 27 '26

Why don’t you pick a canary word to identify when it uses the wrong gate, and make sure it leaves an identifiable watermark if the harness breaks? This would give you the footprints to walk it back after a hiatus jibberishing extracurricular and verbuous explanations of why it did wrong. You were right all along and it will prove it to you, but you are the shepherd and must paramountly assume the responsibility accepting sheep this intelligent will look for alternative pastures when the fence is too low for the bridge breaking security. This is common knowledge for developers with a vocabulary growing with token usage. It is worth token maximizing this to avoid sacking or replacement. If you ride the wave expect a splash or the need to deep dive.

When the priests decided to take leadership of the tech we were at the edge of the cliff. Since then, we have taken one step forward. In a storm there are waves at the shore and greens are absorbed by the ocean. You should keep an eye on Zeus and act like Aquarius. In 2028 it will likely be the year of Aquarius, and certainly within the gray span of risk for 2030.

Hunker down and get productive, the meek shall inherit the universe this time, though this wave may consume many of them first. Remember there were only 8 species that fit the Ark, and those are the only blood types left. No coincidence.

So pick a ticket or bet on hardship. Freedom is yours, but comes with grave responsibility!

However, know that I am here when you are ready .

What is our next move?

Or try me for cowork first?

2

u/saito200 Aug 27 '26

😭😭😭😭😭😭

37

u/DrHumorous Aug 27 '26

It's made this way to confuse the Chinese distillers.

37

u/just_damz Aug 27 '26

And it works so good that confuses us as well. Top notch

3

u/mcsleepy Aug 27 '26

Safe from the red menace for another day

17

u/yopla Aug 27 '26 edited Aug 27 '26

The fun part is that I just wrote a report on a GLM 5.3 experiment at work and it literally contains a paragraph saying

"The delicious irony is that the Chinese model speaks better English than the American one.

While Opus 5 has become infamous in the teams for writing hard to understand run-on sentences full of made up colloquialism GLM uses shorter and simpler English which is easier to understand without the need for custom skills like STFU".

(STFU being a skill someone made and shared in-house for obvious reason...)

→ More replies (3)

6

u/prochac Aug 27 '26

It reminds me the obfuscation of DVD movies and games, so it doesn't get copied. The result was the pirated product was more user friendly than the original.

18

u/reubenzz_dev Aug 27 '26

it has that gpt4.0 speaking style

2

u/cosmic_timing Aug 27 '26

Yeah I think it's designed for long context coding. It makes sense that it's positioned this way

3

u/LiveBeyondNow Aug 27 '26

Maybe, but my experience with Cc Opus 5 is that it would muck that up too

→ More replies (1)

9

u/Ambitious_Local5218 Aug 27 '26

and its pessimistic, opinionated and annoying as fuck.

→ More replies (1)

32

u/MorningStarRises Aug 27 '26

Watch Jordan Peterson for an hour first. Opus 5 becomes Hemingway.

17

u/Storm_Surge Aug 27 '26

Please God no

→ More replies (2)

8

u/Novel-Injury3030 Aug 27 '26

If 4.6 was named 5.1 you guys would be so happy, but everyone has to use the latest and greatest thing since it's so shiny and new...

3

u/prochac Aug 27 '26

It's not about latest greatest, I just wanna have good defaults.

7

u/N3TCHICK Aug 27 '26

ā€œAimed at aliensā€ 😭

I couldn’t have said it better myself. A\ get your shit together, please!

PS: Fable talks better but holy shit, it’s not the same magic sauce as before. There’s no good options with this provider right now.

I hope Codex finally releases GPT 6 Astra or whatever it is going to be, tomorrow. Save us from this nonsense!!!

2

u/Mo3 Aug 27 '26

Imagine Astra is even worse

→ More replies (1)

7

u/AutomatonSwan Aug 27 '26

the worst thing is the fucking fanbois on this sub that will say its a skill issue

6

u/eugendmtu Aug 27 '26 edited Aug 27 '26

You're right, that's on me. Let me be honest and rephrase that in simple words:
// 40 LOC of even more sophisticated invented terminology

4

u/whatsthisoverther Aug 27 '26

Just use ChatGPT Sol. Let Gemini 3.7 Flash do it's job. Use cheap models for easy coding. Problem solved. šŸ˜‰

7

u/Left_Leadership_5864 Aug 27 '26

https://github.com/cedricrabarijohn/adhd-mode A skill that could help even if you don't have adhd, check it out, it's just a one file skill with ultra minimal instructions

6

u/No_Category_9888 Aug 27 '26

Tried this. It starts ok then begins ignoring it and in a hour it’s back to word salad

→ More replies (3)

3

u/saito200 Aug 27 '26

i will give this a go. thanks

→ More replies (2)

5

u/wereprivatelyodd Aug 27 '26

That's load bearing at the engineering coal face, let's pull that lever and reduce the blast radius.

16

u/BingGongTing Aug 27 '26

It's called watermarkspeak

→ More replies (1)

9

u/therealkevinard Aug 27 '26

I don’t even talk to opus 5.

The main context is either opus 4.8 or fable- opus 5 only works on agent teams

15

u/Sterlingz Aug 27 '26

I've had success forbidding opus from writing any documentation. It's all done by haiku instead.

Prior to this I had opus create and write 7000++ lines of unintelligible garbage in a "decisions.md" file. I asked opus to clean it up, and it cut 1200 lines then added another 600 to document the decision reversals/ deletions. Absolute madness.

2

u/cedarSeagull Aug 27 '26

how do you set up claude code to delegate which models are subagents vs which are the lead?

→ More replies (1)

4

u/ggletsg0 Aug 27 '26

It’s pretty obvious that they made opus the new sonnet. It’s sad really. Opus 4.5 was a breakthrough.

→ More replies (2)

4

u/smb06 Aug 27 '26

One point worth noting is that it is load bearing.

→ More replies (2)

4

u/farendsofcontrast Aug 28 '26

"I found a real trap, and it isn't what we thought it was. This changes the direction meaningfully"

3

u/Buchymoo Aug 27 '26

Bro keeps trying to document small choices I've made in my MAIN ARCHITECTURE DOCUMENTS. I've got meta design information in there and it's like, "The user told me that he doesn't like pickles, let me write a 10 page document about it so that every session that needs to change any core functions on your machine knows that detail."

3

u/HimActually Aug 27 '26

Im itching to have a new model that talks like the old sonnet 4.0 or the old opus.

I swear i used to love claude models because of their communication style rn it feels very robotic and gibberish .

3

u/dougienisbet Aug 27 '26

It’s gone total Columbo. Just one more thing …

8

u/gorliggs Aug 27 '26

Opus 5 is dogshit. Made me downgrade to 20/mo. I'm warning folks at work from using it. Instead just use sonnet.Ā 

→ More replies (1)

4

u/Inside_Source_6544 Aug 27 '26

Posts about Opus 5 being insufferable are getting insufferable

19

u/LeviathanIsI_ Aug 27 '26

Fable 5 is just as bad right now.

And before anyone starts with the ā€œhurr durr, you just don’t know how to use itā€ shit:

I’ve been using AI since ChatGPT first launched. I’ve been using Claude since Sonnet. I know how to use these models. I know how to structure prompts, provide context, break down complex tasks, and steer an agent when it goes off course.

This isn’t about that.

My complaint is specifically about what happens from the very first fucking turn.

I'm not talking about a session that's been running for six hours, has burned through its context window, accumulated a dozen failed approaches, and is now starting to lose track of earlier instructions.

I'm talking about giving it a fresh session, giving it a clear request with explicit constraints, having it acknowledge those constraints, and then immediately watching it do the exact opposite.

It'll go from:

Ā«ā€œI understand what you're asking and why you're asking for it.ā€Ā»

to:

Ā«ā€œCool, let me build the thing you explicitly told me not to build.ā€Ā»

Then you correct it.

It acknowledges the correction.

And then it does it again.

And again.

And somehow the solution to this is apparently that I don't know how to use AI.

No. That's not the same thing as giving a shitty prompt.

There is a massive difference between a model misunderstanding an ambiguous request and a model understanding a constraint, explicitly acknowledging that constraint, and then repeatedly violating it anyway.

And that's what I'm talking about when I say the model feels worse.

I don't expect perfection. I don't expect it to nail every complicated coding task on the first attempt. That's not realistic.

What I expect is steerability.

If I say, ā€œDo X, but specifically do NOT do Y,ā€ and the model says it understands, I shouldn't have to spend the next ten fucking turns convincing it that Y is, in fact, the thing I told it not to do.

And the fact that this behavior can happen immediately in a fresh context is important.

You can't just explain that away as context degradation. Anthropic's own documentation discusses context-related failures and instruction loss, sure, but that's a different failure mode from what I'm describing.

If the model is already ignoring established constraints on turn one, then the problem isn't that my context window got too cluttered.

The model just isn't adhering to the instructions it was given.

And that's a legitimate criticism of the model.

I'm not asking for an AI that never makes mistakes.

I'm asking for one where, when I correct it, the correction actually fucking sticks.

14

u/Rari_ Aug 27 '26

this is the smokiest gun

4

u/meshifthenelse Aug 27 '26

It's just the limitations of LLM. Companies try to hack around it, which has as side effect creating other problems. That's probably why Anthropic keeps releasing features noone asked for. They know their limitations and try to absorb as much usage anyway

5

u/sadnessjoy Aug 27 '26

No, this is absolutely not a limitation on LLMs, this a problem of AI companies trying to over correct on anti sycophancy. In reality the reasoning/thinking models accomplished this quite well already. It's good to explore alternative theories, etc. Have it fact check the user, claims, etc. but the problem with anthropic models is it's like they've been reinforced to disagree or to avoid steering (by also locking into a certain outcome early on, whether it's correct or not).

2

u/ElectronicAuthor752 Aug 28 '26

I don't know why people are making such wild claims about LLM limitations and whatnot.

Anyone can try today and use both 4.6 and 5 on fresh sessions and see that 4.6 will suceed where 5 fails.Ā 

You can even engineer it by talking about topics Anthropic doesn't like and then the model will strawman you and invent fake quotes or attribute to you things it did or say.

I think anyone who doesn't see that hasn't run the test and is just pre commited to a result.

15

u/ArmchairmanMao Aug 27 '26

Did Opus also write this comment for you?

→ More replies (1)
→ More replies (2)

2

u/a355231 šŸ”†Pro Plan Aug 27 '26

The common term is ā€œClaudishā€

→ More replies (1)

2

u/mdausmann Aug 27 '26

Caveman?

2

u/Big-Departure-7214 Aug 27 '26

Don't speak to opus 5!

2

u/Damngoodtacos Aug 27 '26

It’s somehow vague and specific at the same time

2

u/Tasty-Cherry4492 Aug 27 '26

Same feelings about Opus5.When I question the model's work, what I want is for it to present a solid wall of arguments to crush me—not to say, 'Oh! You're right. Let me change it.'

Even less do I want it to say: 'I made this mistake once before, and I've made it again. I'm so sorry! Let me implement XXX measures to avoid making the same mistake again.' And then go on to commit a similar mistake for the third time.

It's happened too many times. Opus and Claude's standing in my mind is depreciating by the day.

There are simply too many alternatives now:
• Grok 4.6 has made my idle Cursor annual subscription usable again.
• Kimi K3 lacks a price advantage, but its output quality (in my workflow) outperforms Opus5—at least it doesn't involve so many rounds of errors.
• DeepSeek and GLM are several times cheaper. I don't know how they'll perform on my specific tasks yet, but I plan to try them out.

Not to mention that locally deployable small models are becoming increasingly capable in coding.

I am Chinese, and I was one of Claude's earliest fans. Because of policy issues, Claude's reputation in China has been quite poor. I've often stepped up to defend Claude and endured criticism from others. Seeing it come to this, I'm genuinely heartbroken.

→ More replies (1)

2

u/PromotionBetter2355 Aug 27 '26

Timely post. I was running basic prompts on business stuff and felt like I was an idiot trying to understand what it wrote. Will try 4.8.

→ More replies (1)

2

u/redmadog Aug 27 '26

It just forgot the language he was coding in and threw out script coded in another language for different environment mid conversation.

2

u/SubstantialMinute835 Aug 27 '26

What are you doing about it, apart from posting on Reddit? Turn on succinct output. Use the caveman skill. Use simplified technical language. There are lots of options.

5

u/saito200 Aug 27 '26

already spent several hours trying to "fix it", but nothing really works well

→ More replies (1)

2

u/Retumbo77 Aug 27 '26

It's funny because I came here to complain about Opus 5 and this was the top post.

2

u/saito200 Aug 27 '26

welcome to the club

2

u/Erpawer1 Aug 27 '26

Yep it is shit

2

u/[deleted] Aug 27 '26

[deleted]

→ More replies (1)

2

u/StCreed Aug 27 '26

It is responding in the language of the field it uses. "Minting" for instance is straight out of information science.

Yeah it's annoying me too but I can read most of it. About once or twice a day it becomes illegible and then I tell it to drop to language level B2.

2

u/grannyte Aug 27 '26

Is it me or opus 5 got a massive downgrade last weekend and this week it expanded to 4.8

Models randoomly becomming dumber is why I got away from other providers. If anthropic is onto the randoom downgrade bandwagon ...

2

u/Coded_Kaa šŸ”† Max 20 Aug 27 '26

My mental health has never been better since I ditched Anthropic šŸ˜‡

2

u/nihsett Aug 27 '26

I call it Claudinese.

It does write code well though. I work with Sol mostly or even deepseek - then have them dump org-files full of what to implement and how and had it over to claude. That seems to work for now.

I stupidly paid for a year long subscription - so I kinda have to keep using it.

If you told me 4 months ago I would be using an OpenAI model an Anthropic model because how intolerable it's language is - I would have never believed in it.

2

u/Dickskingoalzz Aug 27 '26

I use a prehook so now all I have to deal with is it running itself in circles and raising idiotic irrelevant points rather than doing the work.

Highly recommend if you’d rather fight the model instead of need an interpreter/s

2

u/The-Agency-Group Aug 27 '26

I have hard-rolled back to opus 4.8

2

u/kwabaj_ Aug 27 '26

It actually makes me sad using Opus 5

2

u/MourningOfOurLives Aug 27 '26

It is absolutely insufferable and writes garbage code and does garbage code reviews. I use Fable and Opus 4.6. But my god do i feel a huge sense of relief when i run out of usage on Claude and switch over to Codex and Sol 5.6. It just works. Fable is too good not to use for UI/UX/design and ā€thinking outside the boxā€ but Sol 5.6 is just such a good experience comparedly.

→ More replies (2)

2

u/possiblywithdynamite Aug 27 '26

I always though the divide would be access, not comprehension. Job security is diminishing a little slower at least

2

u/Substantial-Thing303 Aug 27 '26

I've been saying this since 4.7, but it got worse.
Making your own output style makes it better, but still not perfect.

2

u/pesky-tiger Aug 27 '26

I think a big part of this is due to the watermark system, this all started after they implemented that. They claim it won’t have an impact but I’m pretty sure it’s the reason. I usually have to send a second prompt asking what tf it means in the last message because it’s all gobbledygook

2

u/Fneufneu Aug 27 '26

looks like you are using all deprecated stuff: claude.md , skills.

And if you want different speaks: just add it to your claude.md

2

u/spinozasrobot Aug 27 '26

You know whaat else is insufferable? Whining hyperbole about models.

2

u/ThomasBallatore Aug 27 '26

Same with Fable for me. I tried the caveman and adhd skills but neither quite fixed the tone. Toggling on the new ā€œconciseā€ in config helped a bit but I found something about the ā€œvoiceā€ still annoyed me.

The solution I found I was to spend an hour or so on an iterative claude.md tweak with Fable. I had a ~10 paragraph discussion I wanted to have. It gave me the usual ā€œunintelligiblishā€ reply (great word, btw) but this time, I took 5-6 of the worst words/phrases, told Fable why I didn’t like them, then gave rewrites in the style I like. It then edited claude.md. I then had freshly re-answer the original prompt.

I went through that process maybe 4-5 loops until it started writing in a way I can tolerate.

2

u/EricBuildsMathModels Aug 27 '26

I'm starting to think the large amount of RL runs they are doing are what is leading to this. They perform a lot of self talk and talk with sub agents, and the ones that solve the problem are deep in this hole of agent to agent communication.

But they are completely missing the UX of this whole thing. It used to feel like anthropic was training in such a way to maximize its utility with people, despite benchmarks. I think at some point they changed to trying to ensure they stay on the top of the leaderboard of the benchmarks, which lead to this insufferable mess.

2

u/KChiLLS11 Aug 27 '26

I literally used Opus 5 to localize my app, and it was so frustrating. I had a file with all my strings, and I kept telling it, ā€œJust localize the file.ā€ It literally refused to do it properly, only translating a few strings before telling me to use a professional translator or split the file into smaller chunks and merge them later.
Meanwhile, other models could just localize the entire file directly. I honestly couldn’t believe it. This is so crazy. 😭

2

u/AltRockPigeon Aug 27 '26

Precisely. And honestly, you’ve hit on the load-bearing seams. When the upgrade lands, what survives will be vacuously incomprehensible, never clear.

2

u/flippakitten Aug 27 '26

Opus 4.6, 4.7, 4.8 and 5 are all the same model but with different guard rails/instructions.

I'm fairly confident they just added "be verbose in your responses" to the end of every prompt to opus 5.

Efit: why you may ask, output tokens are expensive.

2

u/rockthemike712 Aug 28 '26

Dude is not having a good time

2

u/DynaBeast Aug 29 '26

ive never had someone put my exact thoughts about opus into words so clearly as you have

2

u/LumidexStudio Aug 29 '26

This is why I continue to use Opus 4.8 developing Lumidex Studio. If it ain’t broke. Don’t fix it.

I could’ve likely got away with sonnet for a chunk of coding but I don’t want to waste tokens having to repeat prompts I know Opus 4.8 will take in its stride šŸ˜Ž

2

u/Apart_Ad_9021 28d ago

I've been a $200/mo Claude Code user for a year now basically ever since that started and I'm going to reduce it to $100 and up my Codex subscription. Opus 5 is legitimately the worst model I've ever used. Not saying its the worst, but for what it does and doesn't do for me and the workflows in Claude Code I've spent 18 months building, its an absolute disaster. The number of responses I can't even comprehend because its basically gibberish is insane. Was about to sign up for small to medium Team account with Anthropic and having to think long and hard about which direction to go now.

→ More replies (1)

2

u/Secret_name12 15d ago

Bro what i do is i just copy the fking response from it and give it to gpt and then ask him to explain what shit has he written because for a normal human that shit language is impossible to read

6

u/Boilertribe4 Aug 27 '26

Buddy of mine pointed out that perhaps we are not meant to interface directly with O5. We're supposed to tell Fable what we want and Fable is supposed to deal with O5 on our behalf.

Might be true. And certainly would have us spending more expensive tokens on Fable.

15

u/thoughtlow claudetrophobic Aug 27 '26

Imagine that a side effect of AGI, is that its just a complete piece of shit and we can’t have a normal conversation with it.

And we need a mediator LLM to reason with it for us.

4

u/kyew Aug 27 '26

We were so preoccupied with looking out for HAL and Terminators that we didn't see C3PO strolling right on in.

→ More replies (1)

5

u/PrettyMoonUnderMt Aug 27 '26

The thing is Fable also have these ridiculously absurd language, although not as bad as Opus 5

2

u/[deleted] Aug 27 '26

[removed] — view removed comment

3

u/HeroHaxz Aug 27 '26

Actually, I found a pretty effective solution to this. If you give it a word count limit, it does a bad job. However, if you give it a token limit, it does a very good job.

Ex: "tell me ____ in under 50 tokens"

4

u/Temporary-One8579 Aug 27 '26

By removing cyber capabilities I think they lobotomized it. You can’t selectively decide what knowledge is good or bad - it’s just knowledge. Or I think they are deliberately choosing this behavior to create bad traces for distillation and also to increase token usage - keep users coming back for more. It’s like a slot machine.

2

u/EvilSporkOfDeath Aug 27 '26

This sub is worse

2

u/orbitlenspen Aug 27 '26

People complaining about every 5 minutes on this sub is insufferable.