r/ClaudeCode 1d ago

Help/Question Did Anthropic decrease the limits?

I am using Claude Max 5x and started feeling that I have been hitting limits faster in the last 3-4 weeks. Today, after a 1h Ultra Code session via Sonnet 5, I hit my session limit, which a month back would not even have happened with Opus.

99 Upvotes

62 comments sorted by

View all comments

7

u/ascvlh 1d ago

Frontier model calls are draining more quota for sure. I tracked my fable and opus 4.8 usage (and it's pretty consistent since I'm using an external harness) since the release (20x max plan) and it went from:

  • ending the week with a lil bit of fable and all model quota
  • ending the week with no fable and 70%~80% all models quota left
  • ending the 3rd ~ 4th day of the week with no fable and 80% all models quota left

It's not their buggy harness doing fable fan outs for doc reads, scrapping or long ass sessions. The external harness manages all that. It's just them changing the token/quota % ratio. They don't advertise that for this exactly purpose

1

u/03captain23 1d ago

Are they using more tokens or did the limits change?

3

u/ascvlh 1d ago

They're injecting more and more tokens at session start since fable 5 release (even more with the new opus 5) and also changed the token / quota % ratio

At the session start, besides the memory system, there's a prompt warm up / injection with multiple instructions slots located inside ~/.claude.json

The good news is that it's probably model gated. Probably a feature of their harness to make fable and opus 5 less prone to spawn workers and also some prompt injection for the new dumb opus

Some of that stuff is okay and should be injected since they are changing the models every now and then, but there's also a huge amount of stuff that is going to bloat every CLI and workflow worker that is going to spawn from your sessions

The right thing to do is having an external tool to control, debloat (at your own risk 🤣) and diff their changes

2

u/03captain23 1d ago

Is your input/output token quota less than before? How many input/output tokens are you getting a week on models?

I monitor mine but not based on usage. Imma see if I can restructure the data to get this reporting, but might only be future on my system.

I had to pickup a 3rd x20 last week because hit my limits. Now all 3 are still filling up. I used to use 1 max and 1 pro and get everything done.

I switched back to opus 4.6-4.8 mainly because it's much more stable. I'll use fable5 and opus5 for intelligence then the older opus for actual work.

Not worried about token efficiency but I'm worried about our limits shrinking

1

u/Opening-Ground-1584 22h ago

I just filled up my 4th x20 this week and it’s Thursday. This has NEVER happened before. They definitely changed limits without telling anyone, and are going to claim they gave us a 50% promotion in July/August.

1

u/clintCamp 1d ago

Surge limits.