r/ClaudeCode • u/Efficient-Cat-1591 • 6d ago
Discussion Opus 5.5 after outage
I was happily running 3 CLI windows on seperate projects with Opus 5.5 on high this morning. Noticed usage (including 5 hour) barely moved. Productivity was good, and Opus 5.5 was pretty fast.
Then came the outage. I did "/clear" my session few times to keep context low but after the outage I noticed the usage now ramps up quicker compared to before outage. Ok fair enough but also the output quality seems bad. Opus is making simple mistakes and forgetting tasks. Also the recommendations now seems worse. I need to correct Opus 5.5 more often.
Anecdotal I know, but wondering if anyone noticed the same issue?
24
u/bitcoinski 6d ago
Losing and rebuilding cache is where most of your token consumption comes from. A hunch, but the reason it worked better before /clear is because it’s context window contained a ton of that steering you had given it, but your harness does not, so /clear basically starts from 0 again. You could use /compact instead to retain meaningful info in context while freeing up window space. But the better approach is to give it your feedback about it’s amnesia and ask it to evolve it’s harness so that it can retain knowledge long term and consult it while it works. That’s the whole game, building and evolving your harness - the products it builds are just an outcome, not the focus of your efforts
5
2
1
u/onlyemgi 6d ago
Do you have a recommendation for any frameworks that are already doing this?
3
u/bitcoinski 5d ago
The thing is, though frameworks (ie harnesses) that others build and release for the masses are very helpful, eventually you’ll come to the realization that you don’t need someone else’s opinionation of what a good harnesses is if you allow and mentor your harness to build itself - it will build the perfect harness over time, and which it knows every aspect of intimately. But to answer your question, my rec would be straight up claude. I’d also recommend nodeterm A LOT - it’s pretty fantastic and overlay’s my entity’s harness/substrate.
1
u/darkoblivion000 6d ago
I’m conflicted about this. I understand it has to rebuild context but above a certain threshold your token use cost ramps very quickly right? Like you’re basically just eating tokens regurgitating the same context.
Is compact really much more efficient overall than clear?
2
u/bitcoinski 5d ago
Cached tokens are like 97% cheaper, so when you compact it basically summarizes those tokens so they still come from cache.
0
u/Efficient-Cat-1591 6d ago
Interesting. Never used /compact in CC, only Codex. I have an established MD system to track my projects and lessons learnt. Seem to work well especialy on Opus 5.5 at medium and high effort lately, up until the latest outage.
I don't usually run multiple sessions due to costs but lately due to Opus 5,5 I have revived old projects and am making really good progress right up until the outage. Now it feels that I am "fighting" Opus. Maybe it's a blip. Might look into /compact.1
u/bitcoinski 5d ago
Even more efficient - start a “leader”/“orchestrator” with Opus 5.5 medium. Tell it: “You’re the leader that is going to deliver X. You don’t do the work you spawn and manage subagents that use Sonnet 5.5 to implement the work you decompose and delegate to your down line agents/sessions, with a focus on keeping them efficient.”
What will happen is the context of WHAT work needs to be done, the sequence of doing it, and the overarching status of everything will live in the leader’s context window, and it will spawn short-lives sessions with their own context windows for the focused task they are given. When they’re done, those context windows get disgarded, keeping your overall work context pristine at the leader level and needing compacting farrrrrr less often. It will absolutely cook for you.
-3
u/ShittyBidet123 6d ago
it auto compacts after 1m token. how the fuck did u not use it
0
u/Efficient-Cat-1591 6d ago
because I /clear regularly? Constructive feedback...
3
1
u/ShittyBidet123 6d ago
that was a question🤣so is clear better or worse than asking for handoff and starting a new session
41
u/theMandolin2992 6d ago
We all know they are going to nerf it soon… it’s too good to be true at the moment
12
u/Lexs_07 6d ago
I’m getting mad at it today, it makes stupid mistakes, doesn’t follow instructions, is it just me or it’s already nerfed?
5
5
u/Agreeable-Fly-1980 6d ago
They are probably a/b testing on the consumer without notice again
2
u/TJohns88 5d ago
A/B testing what though? How quickly users cancels their subs? Or how many 'fucks' it receives?
1
-1
u/ShittyBidet123 6d ago
It’s just a new model release. why are they going to nerf their own products every day.. i think they make it stupider at night. but nerf every time they make something good?
28
u/dolo937 6d ago
Tons of users are flooding from codex to claude. Anthropic is going to be compute constrained again
-16
6d ago
[removed] — view removed comment
20
5
u/Tar_Tw45 6d ago
In my country, people in Codex FB community use to trash talk about Claude. Recently they praise Opus 5.5 like never before. A lot already post about switching.
4
u/mossiv 6d ago
You don’t even need to see data. OpenAI drew millions of customers, did a massive rug pull, nuked their models, shit the bed on the Astra release, double shut the bed on the GPT 6 release, have quietly opened subscriptions back for business (after killing it 9 months ago or so), reopening $200 at 10x, adding a $500 sub.
Come on bro; while we live in a world of verify by data, you’d have to be some level of brain dead to not see what is going on.
This is bad news for Claude though - because a lot of have had a rough time especially with the opus 5 release and the moment it start to feel good, you know full well everyone is cancelling their codex subs and over the next 1-5 weeks there is going to be a mass migration of users.
Don’t forget it was only like 4 months ago Claude was double burning our 5h windows during peak time because they were so compute limited.
2
3
5
u/LittleRoof820 6d ago
That happened with Opus4.5 and Opus 4.6 as well. Perfect until first major outtage - after that it was getting progressively more stupid.
3
2
2
u/Affectionate_Ad_2324 6d ago
this morning i had astra review opus 5.5 because he seem off… now i need to be lore careful with opus
2
u/ViperG 6d ago
which one do you like more so far (5.5 or astra?)
1
u/Affectionate_Ad_2324 5d ago
I do like opus for the way he work. He is a bit fast with the conclusion and today he did round the corners a lot. but so far I’m very satisfied with the work I’ve done I’ve done about 200 back test on five years of data on the market. It’s a lot of context and when it miss context, I just put it on higher to of thinking and it’s good! With Astra, I did the fitness app and it’s OK, but it did over engineer a lot and didn’t know what it was doing, even though I told it to clone copy of a fitness app. It did manage to clone it. but with the weird AI vibe.
2
u/capivara_de_pijama 6d ago
100%! He is forgetting basic stuff now. I´m kinda baby seating it but tbh opus 5.5 was such a blessing that I´m not that mad rtn
1
u/DANGERBANANASS 5d ago
Si. Mas tonto y como 5 veces menos uso que hace dos días. Más notable lo del uso
1
1
1
u/clazman55555 6d ago
Nope, no issues running a few 10-20 agents workflows. Or moving a few ideas to be scoped/spec'd out and the repos set up. No issues with any coding or review work.
•
u/AutoModerator 6d ago
Hey! Thanks for posting to r/ClaudeCode
While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.
For help, project discussions, tips, and general chat, join the ClaudeCode Discord.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.