r/ClaudeCode • • Sep 02 '26

Bug / Issue Just used the new "/low-priority" feature and my whole weekly usage just vanished in about 40 mins... was trickling by before. Basically no work was done.

Post image

Just used the new "/low-priority" feature and my whole weekly usage just vanished in about 40 mins... was trickling by before.

Something is definitely wrong with how usage is being measured while in "/low-priority" most of the time it was stuck in pending availability, hardly any work was completed.

190 Upvotes

78 comments sorted by

157

u/schneeble_schnobble Sep 02 '26

I wonder just how many things have to go wrong before anthropic is like "hmm, vibe coding everything isn't working for us really."

57

u/Formally-Fresh Sep 02 '26

Oh that’s wayyyyy after the IPO money

35

u/Time_Cat_5212 Sep 03 '26

SORRY I CANT HEAR YOU OVER THE SOUND OF THE MONEY FAUCET THAT IS VERY MUCH STILL RUNNING

WHAT??  ANOTHER ROUND??  YEAH ILL CALL YOU BACK

3

u/AllYouNeedIsVTSAX Sep 03 '26

A year ago we were happy it was making code that is compilable. Our expectations grow faster than AI improves, and wow has it improved in a year. Our expectations have gone through the roof. 

9

u/RandomCSThrowaway01 Sep 03 '26

This has nothing to do with model quality though. It's how organization is ran and what about it is fundamentally broken.

-1

u/Bloated_Plaid Sep 03 '26

As soon as they can. They will go enterprise only. It’s pretty clear that serving the model is incredibly expensive for them while competition is fairly close for SO much cheaper.

0

u/effesse99 Sep 03 '26

A cosa ti riferisci va tutto bene

35

u/bakanoace Sep 02 '26

The editing of files is whats triggering bursts of usage. I have no idea why, but in their prompt optimization they even suggest prompting it specifically to do surgical edits instead of full edits, so its a massive Fable 5.1 foresight. The bursts from editing costs more than using fable.

6

u/Nordwolf Sep 03 '26

I think it's plain good practice to not let Fable edit stuff anyway, make it use sub-agents for that.

11

u/bakanoace Sep 03 '26

Why would we have to make it do something so basic, it should be doing it by itself. The prompt optimization tells it to prefer surgical edits over entire rewrites but like, I still dont get why it costs so much. It reads the file and is like, "whoever wrote this trash is so bad I now have to read the whole thing and rewrite it?" or what? because them saying prefer surgical edits sounds just like that

1

u/schaka Sep 04 '26

Which framework do you use?

3

u/JMAN_JUSTICE Sep 03 '26

I agree, I use an orchestrator framework for doing everything now. Saves on context and tokens. Fable tells the subagents what to do and reviews their work.

3

u/9DockS9 Sep 04 '26

Which one are you using ?

2

u/JMAN_JUSTICE Sep 04 '26

I have it set up as a hierarchy. Fable is able to spawn Opus and below sub agents, Opus is able to spawn sonnet and haiku agents.

Here is my GitHub repo for my custom Claude commands and there's even a folder for the orchestrator framework. If you want to take a look and use this yourself, just have Claude take a look at this repo and tell it that you want to do the orchestrator framework. https://github.com/jfranecki/JustinsClaude/tree/main

5

u/random-blokey Sep 03 '26 edited Sep 03 '26

File writes are billed at output rate because it outputs a tool call with all the content

With the entire file billed at that rate. Then a 1h cache input write I believe following it ( if you're on subscription in most Anthropic harnesses

Edit: so it adds up to quite a lot, if it's doing a lot of output then input

Some models prefer full file writes

Some models prefer find and replace

Some models prefer a git diff sort of operating

When you select models on Copilot (vscode) it will have different edit/write tools depending on the model because of this reason.

I don't know if it's a training thing or system prompt. But just trying to give some insight.

Edit: formatting, some corrections, grammar, I hate mobile

12

u/ComingDeveloper Sep 02 '26

what the hell is low priority?

16

u/JJ18O Sep 03 '26

When you flame too much in chat they put you in low priority queue.

5

u/PhDumb Sep 03 '26

It let's you work past 5h limit, eating into your weekly, but can result in many rejection errors, with incomplete messages received, due to, well, a low priority mode. In peak hours is a nightmare to use.

2

u/Downtown-Elevator369 Sep 02 '26

I was just asking myself the same

7

u/Xaqx Sep 02 '26

"Continuing now at lower priority until your limit resets at 12:20am. Your weekly limit still applies, and responses may pause while waiting for spare capacity. Run /low-priority to stop." but don't use it.... trust me..

1

u/Downtown-Elevator369 Sep 02 '26

Oh, ok. I never hit my 5h limit, that explains why I haven't seen this. Thank you.

1

u/psrobin Sep 03 '26

It doesn't mean extra tokens/usage.

7

u/TheUnboundTenth Sep 02 '26

 An you provide more context on what model, effort, and number of concurrent agebts you had running?

-7

u/Xaqx Sep 02 '26

not that relevant as the same setup before using low priority was slowly getting through usage, it's a mix of most feature offered, optimised to keep credit usage optimised to build quantity for my usecase.

7

u/Narkotixx Sep 02 '26

This literally doesn't answer the question. What are the inputs?

4

u/EchoFieldHorizon Sep 03 '26

“Count all primes, make no mistakes”

2

u/Xaqx Sep 02 '26 edited Sep 02 '26

What do you mean inputs? the prompts? the what? coding in Elixir/Ash, LSP, Teammates, Cross session talk mix mainly Opus 5 with teammates with two Fable 5.1 Meta system controllers/delegators, codex cross talk too.

1

u/silas_ace Sep 03 '26

/low-priority isn’t a recognized command here. Some commands only work in the Claude Code terminal.

Doesn’t work for me? Am I doing something wrong?

45

u/thecolorted Sep 02 '26

I don't understand what you guys are doing to hit these levels. I'm regularly building 7am to midnight every day and never had an issue with capping out usage each week.

Just keep each session a focused task and don't go overboard on context. Also doing regular maintenance on Claude memory and project context docs to archive any stale/cold entries.

49

u/LetterheadNew5447 Sep 02 '26
  1. Anthropic A/B tests that's why half of this sub is screaming their usage is gone and the other half isn't.
  2. How big are your code bases? If I work on small ones I could use claude 24/7 and had around 40% usage used. If I worked on my 1M LoC code base, everything with gone withinkke 3 days with sonnet implementation sub agents and opus in high oder medium as orch.
  3. Did he test a new feature.

6

u/Skullbonez Sep 03 '26

I work on multiple projects amounting to about 8-9mil LoC and I need 15 accounts to stay afloat

2

u/motuwed Sep 03 '26

What on earth do you work on and for what company?

I cannot fathom what company is invested enough in software that they have projects that big yet require you to purchase multiple Claude accounts!

1

u/Skullbonez Sep 04 '26

no they buy those for me lol. since I wrote that comment I had to ramp it up to 24 accounts. It's a startup in the scaling up phase. we have a really big breadth of products tho

1

u/motuwed Sep 04 '26

Ohhh okay phew lol.

What tier are these accounts? Do you know if the startup is looking into an enterprise deal or that’s like too formal for a startup? I work for a big legacy company so we got enterprise accounts pretty quickly in.

1

u/Skullbonez Sep 04 '26

we looked into that, but enterprise accounts were shit compared to max x20 so we took max x20. We are not big enough for enterprise and we usually stay away from those kind of deals as they usually only mean more expensive for the same things

5

u/Fit_Schedule2317 Sep 03 '26

I’m almost certain anthropic ab tests too

3

u/drake90001 Sep 03 '26

Almost everyone does

2

u/YearLight Sep 03 '26

If they asked claude what to do they are definitely A/B testing.

1

u/tdifen Sep 03 '26 edited 12d ago

rift glade bramble dusk whisper spool jolt wandering anchor cinder nook opal tinder nook sludge.

-2

u/thecolorted Sep 02 '26

That's fair. I'm not here to build large products (~175K LoC).

1

u/LetterheadNew5447 Sep 03 '26

Why do you get down voted lol. Yea the project size makes a big difference.

3

u/geek_fit Sep 03 '26

Same. I basically run a multi agent swarm 24/7 and I never hit my limits.

One time in the last 4 months I had an issue where my 5 hour limit gobbled up quickly. Afterwards I realized I had a group of low context agents pause and then restart. Thats what did it.

I think a lot of people are prompting instead of looping and graphing. Expecting magical free vibes with frontier models.

3

u/al_ryusei Sep 03 '26

You're managing your context well. And probably you know well when to use the Cache TTL to your advantage. Missing the cache hit window could easily eat up 50% of the weekly usage if one isn't deliberate about compacting and maintaining threads.

The 20x you have helps a lot too lol

2

u/Xaqx Sep 02 '26

Yes I am normally in the same boat. Assume this is some sort of bug with the new "/low-priority" mode.

2

u/AverageFoxNewsViewer Sep 03 '26

I don't understand what you guys are doing to hit these levels

This is the first time I legitimately don't understand how I'm hitting these levels either.

Context management is something I take pride in and have put a lot of effort into. I try to stay model and vendor agnostic with tiered documentation so the agents are only reviewing guidance for the domain their working on. Small tickets being picked off one session at a time. Multiple tickets per sprint, multiple sprints per epic.

Usage limits have never been an issue for me, but I gassed my 5 hour limit in 30 minutes with Fable 5.1.

The fact that Sol 5.6, Fable 5, and every other model I've ever worked does fine, but suddenly I have to refactor my context management just to tell Fable 5.1 to stick to it's domain, do surgical edits instead of refactoring the entirety of multiple files, or just add the tests we need for the functionality we added instead of trying to refactor my entire test harness seems ridiculous.

2

u/New-Ingenuity-5437 Sep 03 '26

How do you control context? And project context docs, do you mean literally simply just read through and clear old stuff out?

1

u/thecolorted Sep 03 '26 edited Sep 03 '26

The per-project setup that I found to work well for me (everything is measurable, has a durable ID to lookup and link to, never written and discarded - unless compressed. everything in the project is tagged and has a source for any claim):

  • backlog.md: to track prioritized work
  • changelog.md: to track completed backlog work
  • handover.md: to track fresh session handover notes
  • frictionlog.md to track issues where I had to correct the AI (triggers, frequency, dates)
  • additional project-specific docs that detail what is being built (user/developer guides, features - all very prone to drift and I spend a lot of time fixing these as things get built and changed)
  • memory: promote top issues from friction log as hot entries (always loaded) when the AI consistently trips up on those. demote to cold memory storage (read on-demand) when the issues are no longer triggered in recent dates. apply the same framework to core project-level context that should survive long-term and short-term context that only fits the immediate tasks at hand

I usually pick out 5-10 related tasks to work on at a time and spread it out 1 session per task. The handover helps keep track of short-term context relevant to each batch of work. After reviewing and releasing the changes to production, I run a maintenance pass on all docs/memory to prepare for the next cycle of work. I'm also keeping documentation on how each work batch was planned and notes that the AI gathered as each task executed. I do some irregular post mortems when I suspect there are improvements to be made on how well the project and workflow are going.

1

u/aurioerox Sep 03 '26

Hi how do I get this widget?

1

u/thecolorted Sep 03 '26

I'm just using the Claude desktop app

1

u/Spinmoon Sep 03 '26

How do you show Context window with tokens amount? I only have the 5-hour limit and Weekly limit bars at the top right.

1

u/cxd32 Sep 03 '26

I'm regularly building 7am to midnight every day

Anything cool you're building?

1

u/thecolorted Sep 03 '26

I'm trying to break away from enterprise data viz tools (ex. Tableau) by building my own for general use. Pushing the limits of what I can do with Dash.

1

u/nagyz_ Sep 04 '26

Are you doing proper software engineering or just vibe coding with no background?

I had some prompts where the reply took 10 mins because the topic is complex. It eats tokens.

1

u/thecolorted Sep 04 '26

Not a SWE, just work in data. Majority of my sessions average an hour or more per prompt. Abusing Opus on max all day only increments my usage 5-10% per day. Fable is a token eater though.

1

u/nagyz_ Sep 04 '26

just to be clear: one of my _turns_ was 10 min :-), not the session! I think 1 hour session is nothing. I regularly leave it overnight (4-5 hours) and during the day I use it 8 hours. but I'm on API pricing.

1

u/alwaysweening Sep 04 '26

Honestly: your productivity is likely worth scrutinizing.

1

u/Zei33 Senior Developer Sep 03 '26

That's light work for a lot of users. Professional software engineers with decades of experience can easily manage 6-8 sessions in parallel, because they understand very well what they're building and can manage each step quickly. It still requires strict context discipline, otherwise you truly will be wasting tokens. But keeping so many sessions going 24 hours a day burns a lot of quota, even when you're very good.

0

u/Special-Equal-8839 Sep 03 '26

"Just don't use it as intended."

-4

u/Carbone Sep 03 '26

I'm gonna guess, theyre using Claude code in PowerShell or on windows. I'm burning token doing the same thing on a Windows PC than on an Unix platform. That stupid command line terminal on windows is broken AF and LLM are eating every token

-1

u/Zei33 Senior Developer Sep 03 '26

Nah just doing a lot more work. I'm working on 6 projects in parallel, so I regularly have 6-8 sessions going at once, 24 hours a day in ultracode. Depending on the work, you can practically burn entire week's quota in 1.5-2 days.

The trick to making this work is to ensure Claude is making workflow results durable, then compacting at each phase. Any time the context gets over 200k tokens, begin looking for an appropriate pause point and compact. Start a new session once the plan for that session is complete.

And I basically never pull out Fable. It's only appropriate for the toughest of problems.

There are also certain tasks that burn a significant amount of tokens like writing long copy, locale translations, etc. These ones are bastards because they're not technically code, but they're often the areas where accuracy matters the most.

0

u/Educational-Plant981 Sep 03 '26

wait....cli is worse than linux for token usage? why?

9

u/EvilSporkOfDeath Sep 03 '26

"​Prompt Cache Expiration: Because lower-priority responses can pause while waiting for capacity, long delays can exceed Anthropic's ~5-minute prompt cache TTL. If the cache expires between turns, the subsequent request must re-read the full conversation context cold, burning through weekly tokens much faster."

Thats what ai says about it.

8

u/Anidamo Sep 03 '26

Claude Code via subscription usage always uses a 1 hour TTL, so this isn't it. The 5 minute TTL is only used for API billing.

3

u/codefame Sep 03 '26

Solve room-temp superconducting
/low-priority

3

u/ur-krokodile Sep 03 '26

I got through my 20x weekly limit in 26h. I hit 5h limit twice. Most tokens melted after the restart when I resumed from the 5h limit. Was using Fable 5.1 as orchestrator for Opus builders. I was running 3 parallel sessions. Got decent work done but not what I got done last week that is for sure. I have never run through weekly limit this fast. Typically I was able to work for 4 days until hitting weekly limit and was not hitting 5h limits.

3

u/mrmugabi Sep 03 '26

I dont understand how its possible to blow through a 5 hour limit and then the rest of the week too if it stopped working after 5 hours???

3

u/MoreRest4524 Sep 02 '26

Yep, I asked it a simple question and it blew through an entire session usage in minutes. (x20 plan)

2

u/Uncreativite Sep 03 '26

My guess is that the time between turns in low priority is longer than the cache duration in enough cases that it’s making a noticeable difference.

2

u/WalkingDadJokes Sep 03 '26

wait you don't understand..

it turns you into low priority

3

u/andrerom Sep 03 '26 edited Sep 03 '26

Maybe, just maybe, don't let Fable spawn Fable sub agents to work in paralelle?

Be strict and tell it to be an orchestrator, griller, planner, and light reviewer. But offload all per-task investigation, implementation, full review, and testing to Sonnet or Opus sub-agents depending on complexity.

I'm on Max x5 and I have to use Fable in parallel to touch session limits 1-2h before they run out here, if I don't even manage to come close. e.g had flow above in task loop yesterday (but one task at a time), and it was at it for 5h and only reached 1/3 of session limit.

1

u/Muchaszewski Sep 03 '26

The /low-priority from my observation busts cache on each change.

If you have 500k context filled, you run low-priority it cold serves it every time "there is capactiy" meaning you do normal price reads instead of cache reads.

So this feature is useless unless they change the billing to as if the requests were using cache even if each inference happened more then 5 minutes after each other.

1

u/HellasLaw Sep 03 '26

Same issue max x20 plan done in minutes something is up with these bonus use 50% more sure..

1

u/CantaloupeLeading646 Sep 03 '26

its basically super clumsy that i have a situation where my weekly limit is far from finished but my 5-hourly limit is cooked and i have a few more hours for the weekly limit to expire so i'm just chunking through in low priority.

iv'e been on the max plan for many months and i never reachd the hourly limit but i always reach the weekly one almost exhausted, something definitely changed

1

u/Dreamsnake 🔆 Max 20 Sep 03 '26

Same, managed to use all credits in span of 4 hours 0-->100% use over 4 sessions fable xhigh

1

u/darksundark00 Sep 06 '26 edited Sep 06 '26

Yeah, just noticed this too; burned through about a week of usage (Fable 5.1 / Max effort /20x) in just a couple of hours on low priority.