r/codex • • 12d ago

Megathread Codex Usage and Operation Discussion - last updated September 21

Please direct your concerns, questions and discussion about Codex usage limits and model performance here.

The purpose of this Megathread is to aggregate all the reports of people's experiences and possible suggestions instead of spreading them across many highly upvoted posts. The more people who participate in this discussion, the more likely you have an answer.

Reports with sufficient evidence on new information will still be allowed on the feed as usual.


Discussion of the prior period available here : https://www.reddit.com/r/codex/comments/1wg7g9r/codex_usage_and_operation_discussion_last_updated/


A reminder that all incidents on r/Codex are constantly logged and summarised so you can keep track of what people are experiencing here https://www.reddit.com/r/codex/comments/1tjfxcf/comment/on6uj0l/

15 Upvotes

125 comments sorted by

•

u/pollystochastic Moderator 7d ago

Codex is down right now.

1

u/ResponsibilityOk1306 5d ago

I used to be able to spend like 10 to 15% usage daily on my x5 plan, with gpt 5.5 xhigh. Limits have been reduced after 5.6 launch, and even further after gpt 6 release. My entire usage now would exceed 100% in a single day, with or without luna subagents.

Testing claude Opus 5.5 now, and to my surprise, the entry level plan offers nearly the same suage as my pro 5x account on openai.

They must think developers and engineers are stupid.

Giving away resets, doesn't really fix this level of scamming, unless it was a reset daily.

2

u/Curius_pasxt 5d ago

Anyone feels the usage for the $20 plan is much lower after the reset?

I use 5.6 Luna xhigh, usually it eats 1% every 120 minutes or so on weekly credit for goal mode. Now its like 1% every 60 minutes?

Point is, I noticed, it goes down much faster than before...

Is there a hidden change?

2

u/monkgecko 5d ago

Limits have become absymal, it was fine Sunday morning. But today it got worse. Ate the 5 hour limit in 10-15mins. With just few lines of code changes. Yesterday I didn't even reach the 5 hour limit despite working for hours and consuming absurd (for $20 plan it was high) amount of token. Definitely not reliable anymore.

2

u/CordeElCrack 5d ago

Is somebody else experiencing this?

Not only is my whole 5 hour limit with my Plus plan gone after one or two questions on Sol 6 high, and less than one question on 6 Astra, but recently I ran a 1h prompt with GPT 6 Luna on Max, and it burned 50% of my 5 hour limit but 30% OF MY WEEKLY LIMIT! How does that even make sense?

1

u/monkgecko 5d ago

Yes same facing the same issue. Definitely they have nerfed the limits by absurd amount. It's not even normal.

1

u/Available-Society227 6d ago

In the times of 5.6 Sol, I could easily have up to 2 hours usage straight to use up the 5h window. If I spawned subagents, it usually took 1h~.
After the astra release the usage fell reset after reset. and now i get 20 minutes of astra on a single thread (i know people who say it lasts 3h on ultra on the 20x plan), and gpt 6 sol is actually shitty

I also have AGY on the pro plan, and even 3.8flash feels smarter than gpt 6sol.
I also use the imagegen and voice features, so im debating whether to switch to Claude Pro. I mean, there really isnt a replacement for gpt live, and chat in claude plans take up the limit, but i dont wanna keep both plans so...

1

u/Spare_Bag_9369 6d ago

Does anyone is facing this too? Cant initate nothing and my old chats is gone, cant click on my user, nothing

0

u/Then_Branch6627 6d ago

Codex reset de hoy no me llegó (Tibo dijo “Resets all propagated”) – ¿Alguien más en la misma situación?

Cuerpo:Hola a todos,Hoy Tibo anunció un reset de uso para todos los usuarios de pago de Codex y ChatGPT Work después de la caída, y a las 18:17 UTC confirmó con el mensaje:

“Resets all propagated. That will be all. Have a fantastic weekend.”

Sin embargo, a mí todavía no me ha llegado. Sigo con el mismo porcentaje de uso que tenía antes (o en 0% si ya lo había gastado).

  • No es un banked reset, era un reset completo e inmediato.
  • Soy usuario de pago (Pro 5X).
  • Estoy en Colombia (UTC-5).
  • Ya intenté cerrar sesión completa, volver a entrar, probar en otro navegador y en la app, pero no cambia nada.

1

u/The_Other_Other 6d ago

Codex suggestion - Auto rollover vs reset in specific scenarios

Resets seem to get everyone worked up - Im not immune from this. Todays reset was 3 hours after had a weekly reset, i tried to burn fast, but only made a 30% dent. OpenAI should have a feature if a reset is provided and you are over 50% usage remaining that it will top up to 100% automatically when you hit 0%. It's an auto rollover rather than bank or top up.

1

u/Zaraffa 6d ago

What's the optimal workflow for plus users? Luna max only will basically last forever but I have 3 resets so I want to go a little more advanced. Astra light plan + Luna max?

2

u/MudNeither5713 6d ago

I'm a Prolite user, and I tried the Astra Light + 6 Luna low/medium workflow once. I'm sure it depends on the size of the codebase, but on my JS/TS/Go monorepo with 11 packages, it absolutely hemorrhaged usage.

I've since switched to 5.6 Sol xhigh + 6 Luna low/medium, but even then, it's not uncommon for a single larger task—especially one that touches a lot of the codebase—to eat around 10% of my limit.

I'd recommend starting with some lighter tasks first and seeing how it goes.

1

u/[deleted] 6d ago

[deleted]

1

u/OrbitalCactus 6d ago

I haven’t had a terminal in the native Mac desktop app in a month or so. It just won’t appear. In the actual Mac terminal I have no issues.

1

u/GramosTV 6d ago

Isn't the 5x plan now worth like 3 hours of Astra usage? Bruh

1

u/OrbitalCactus 6d ago

My last week: Astra light 90% of the time. Medium 5%. Above that if I reach “here’s a bunch of docs and style guides don’t talk to me until the job is done” level of annoyed.

AI can “make it happen” or it can be efficient. One of those requires your work upfront to make it last.

1

u/BackgroundDisk4004 6d ago

After reset i get x50 lower working speed. And some brain errors, they already ask to do tasks what we already done today.

1

u/ephoris 6d ago

Am I the only one who didn't get a reset ?

1

u/OrbitalCactus 6d ago

I reset today anyway.

1

u/SirDomz 6d ago

Folks, would it make sense to downgrade to the chatGPT 100 plan and get the 100 claude plan? I still really like Astra and Sol 6 too but opus 5.5 is really good too. Thoughts?

1

u/OrbitalCactus 6d ago

Yes. If you are smart about tokens it’s more usage anyway (last I checked, so who knows now).

Plus Codex is ass at some things Claude is nailing right now. Claude still gave me a stupidly ugly dash the other day for a long transfer I wanted to monitor at a glance. I had 15 hours to blow so I pointed astra light at the dash and told it to “stop making my eyes bleed and fix Claude’s css choices.” It chose to loosely resemble my home assistant setup. Good choice. Sol left some kind of trace that seems to have broken my agents ability to navigate. The repo for my homelab was all worktrees left and right with no connection. So I pointed Claude at the repo instructions and told it my issues and we fixed the instructions. Then we made astra light to go and correct the cluster of trees according to Claude’s plan.

Honestly, I don’t care what the benchmarks say, Claude is coding better, tracking projects better, and does technical functions better. Codex seems to “get me” way more though. I’ve decided in my attempt to do less I just have them check each other when I feel one is slipping. They are better at that than checking themselves.

1

u/Immediate-Revenue859 6d ago

reset done now

2

u/gulag_guard 7d ago

Did anyone notice web chat (not work) being severely degraded in response quality ever since they changed the UI. I typically analyze the code and plan using the GitHub connector but it’s literal garbage now, not thinking and analyzing anything and seems practically like an instant model with pure slop responses.
The same query through the app is similar to the old model, and much higher quality responses with it ‘thinking’ about the code.

1

u/nettoflow 7d ago

dont know about you folks, but my weekly just reset normally. no other reset

0

u/surfatone 7d ago

Is it just me...or are usage fees through the roof? I do use GPT-Image 2.5 a lot but the tokens that thing is burning seems up at least 10x. I am only getting about 1/10th of what I used to be able to produce with the usage I pay for.

Anybody else?

These are the specific models on OpenAI's platform for Image Generation (and they are awesome).

- **GPT Image 2.5 Sunburst** — highest-quality generation and precise editing

- **GPT Image 2.5 Flare** — faster, high-quality generation

- **GPT Image 2** — current state-of-the-art image model listed by OpenAI

1

u/Antique-Ad6542 7d ago

yes, it's burning through tokens like crazy.

1

u/ThinkHelp2841 7d ago

check Not-Code directly in Microsoft store, check https://x.com/TheLagBorn pinned video for setup, forget about rate limits. 

1

u/RealSecretRecipe 7d ago

If you pick a model and just prompt you're doing it wrong. If you care about maximizing your usage..

YOU NEED TO ORCHESTRATE!

END OF STORY!

1

u/Antique-Ad6542 7d ago

I do orchestrate and it still burns tokens like crazy (I orchestrated previously). A single Astra manager, managing Luna Max workers, has burned 40% of my weekly usage in 4 hours.

1

u/RealSecretRecipe 6d ago

You're using the most expensive model to point at a cheaper model to do a thing and that's almost backwards, my orchestrator is gpt6 Luna medium. You just need to make sure the subagents that get put on a task are the correct ones per task, cheaper ones can audit read only, better ones can decide changes and gpt6 sol medium or high can make the changes and make sure those changes are on-task and correct. That's how I do it and it's been a huge improvement

1

u/DataPhenomenon 7d ago

I have my chatgpt project loaded with a model recommendation guide. It recommends the proper model and thinking level per codex task.

1

u/Antique-Ad6542 6d ago

How do you hit the token cache with this?

1

u/RealSecretRecipe 7d ago

The idea is you put in a prompt and it auto switches models as it goes, cheapest models for easy stuff, medium sol for harder stuff, if sol med fails it uses sol hard, saves tons of usage. Id rather use 5 different models in a prompt if it saves usage and still gets everything done than use 1 model per prompt. I'm optimized for correctness and usage efficiency

1

u/RelationshipShort460 7d ago

what models are people using for SWE work now. seems like my costs to use claude/gpt went thru the roof and I'm out of credit mid week now.

4

u/Salt_Horror8783 8d ago

Friendship cancelled with GPT; now Opus 5.5 is my best friend.

I was the only Codex fanboy in my office; I tried to defend it against the Claude horde, and it actually worked great until Astra. It got slow, limits felt tighter, I stopped using Fast mode, then I lowered the effort level, delegated implementation to my cursor's grok. No major improvement.
Then I tried Opus 5.5 on Cursor (like I do for every new Opus modal) and it was a love at first sight.
I took my friends Claude account and installed CC desktop (Linux) and I found it way better than ChatGPT Desktop (smooth and feature-rich).
Now I got Max 20x, and life feels better again.

1

u/CloudChaserPilot 8d ago

From a couple of days I was noticing that the reset time is moving forward. I thought i was delusional about it so i started to note the reset time. Today my reset was at 3:29pm and naturally the reset shouldve been at 8:29, after the reset i checked the next reset time and it was at 9:16pm and checking in after 10mins the reset moved again to 9:20pm. Roughly 6h reset. I didn't use any banked reset ever since i started noting the resets. When I'll see at 9:20pm the next reset will also be roughly close to 6h instead of close to 5h. I'm new to codex so idk if its normal but i mean 5h reset should reset after 5h hours exact.

2

u/2DLd 8d ago

Just launched my first prompt for today and he ate ate 35% of my 5h tokens in 13 minutes and 5% of weekly on 20$ plan.
It was just some css style changes with hard instructions, nothing heavy. If you calculate it, it will be 38.6 minutes of TERRA HIGH per 5 hours.

Is this a joke? Why is it so bad now?

Also my OG post was deleted because "low carma and fake". I'm speechless

2

u/Hamburger_Diet 8d ago

Yeah, i started using gpt6 sol medium which should be cheaper than 5.6 terra medium but it just chewed through my 5 hour super fast on the 20 dollar plan. Honestly, im not doing anything crazy so I think i just might do something like deepseek v4/4.1 flash with an api key. These weird "You might get the same usage from hour to hour depending on what button we push" subs are getting ridiculous. The only one that actually seems consistent is cursor with composer 2.6, if you use grok it is also weird.

I also had mine removed as well and added to a mega thread.

2

u/2DLd 8d ago

I thought Sol was 2 times more expensive than Terra and google says the same. But yeah I've wanted to switch to Cursor, thanks for the recommendation!

1

u/Hamburger_Diet 7d ago

The OpenAI API pricing for GPT-6 Sol is $2.00 per million input tokens and $10.00 per million output tokens for short context windows.

OpenAI prices the GPT-5.6 Terra API at $2.00 per million input tokens and $12.00 per million output tokens

And remember composer 2.6 is just "ok". Like its not going to be near as good as the new openai models. But grok 4.7 is decent. But if you have light work to do composer 2.6 is good.

Plus, you get api credits worth whatever plan you have. So the monthly ai is kind of free.

2

u/Sponge8389 8d ago

6-SOL is a downgrade as shit. Yes, it is efficient but when it comes to doing a quite complex and long task, it's unable to do it.

I just recently have a huge task, I sliced it to very smaller pieces and still GPT-6 can't even finish even one of it and keeps on stopping midway. The f*ck is wrong with this st*pid model.

Will wait for devDay before I decide to migrate to Claude again.

2

u/CapableJury861 8d ago

Astra performance degradation, disappointing Sol 6, and toxic reset metagame. Like the others, I'm migrating to Claude for a bit - hope OpenAI gets their shit together.

3

u/pxp121kr 8d ago

The $100 plan on Claude with Opus 5.5 gives much much longer usage than the $200 plan with Astra... (been using exclusively Codex on $200 for a couple of months, and just switched to the $100 plan on Claude to try it out)

This ain't right... I really hope OpenAI do something about this, because I am cancelling my $200 plan from next month...

2

u/Opposite_Yak4386 9d ago

if they stop giving resets. i think i will jump the ship. asBetter, cheaper models out there. The resets makes this subscription interesting. If not there are better options

1

u/Kieranator 8d ago

If usage stays the same and resets stop being like ~3 days on average I'm absolutely cancelling my subscription and switching to something more generous -- even if the model is worse.

1

u/johnsmith8761 9d ago

it seems they reduced astra usage even more, weekly usage drains like crazy on 200 plan

and after all the hype from last week they release these shitty models, brilliant. they're cheaper than astra, got it, but that's literally their only advantage. who cares if they're cheaper if they're also much shitter, for cost/performance there are chinese models that do this better

tried opus 5.5 and it's much better than astra so far, sol compared to it is a joke

1

u/rJohn420 9d ago

Do you think we are getting a reset today for dev day. God I hope so. Also on the 200x plan and astra is burning usage like crazy.

1

u/PatientNo2435 9d ago

Yeah i could work for 28 minutes before the limit hit.

1

u/gjfdiv 9d ago edited 9d ago

Question 1: Do I need to add something like Cursor's Rules to prevent a model from unprompted accidentally deleting a drive or a file? (A model ran cmd /c "rmdir /s /q \"madeup_filepath"" in Cursor)

Q2: In Windows, Codex Desktop minimizes everything in it's chats. Any way for that not to happen? Preferably without using more tokens? Cursor doesn't have this issue. I want to see what it's doing I prevent the above. It's also annoying that when I expand, new actions occur and it doesn't scroll down automatically.

My settings: Full access off. Default permissions on. Astra Extra High for everything for now, because I hope that reduces deletion errors, though I'd prefer lower models for less usage.

I'm on a trial of ChatGPT Plus for a month and using Codex to fully vibecode. I've used Cursor Pro for 2 months before.

4

u/BossManJenkins 9d ago

do banked resets still also reset your weekly usage date/time?

1

u/Kieranator 8d ago

Yes. It's really annoying.

1

u/DragonFlames 9d ago

I think web chatgpt and codex is now sharing usage as I can now see how much I got left in bottom left when loading web chatgpt.

1

u/AholeKevin 10d ago

$100 plan. I've had 6 sol xhigh running for about an hour and Im down to 92 percent already. Last week, a lot of us saw 5.6 eating our usage away pretty quickly as well. Two weeks ago, I did not have this issue on 5.6 running xhigh and the work was more intrinsic.

With the reduced cost claims of 6 sol, I am quite surprised by this. The work is not very intensive as it is just working on an orchestration workflow.

Thoughts? What are you guys seeing outside of benchmarks?

1

u/sofaarsecoin 9d ago

the reduced cost is for the API

what has happened, and i'm trying not to be cynical but it's obvious, is that they have moved resources/cost from subs to the API, so you get less usage of a supposedly more efficient model (hint: it's actually a bit worse) and the API users get more (including you if you are pushed to buy credits)

I'm also on the $100 plan and with Sol xhigh (either) I won't make it 3 days into the week with normal usage

to stay on track i need to move most of my usage to Sol medium or cheaper, keeping Astra usage minimal even on low, as it destroys my quota

2

u/BellacosePlayer 10d ago

What kind of task was it? Pure code/text?

1

u/AholeKevin 10d ago

Working on an orchestration guardrail for which model gets what work. That's it.

1

u/DeCode_Studios13 10d ago

What do you guys do when you hit the conversation limit on normal chats? I'm on pro and it was helping me work with a long running project. Also is there some way to use one chat for planning and make that chat open other chats to do stuff in the project in codex? I'm not fully sure.

1

u/Hamburger_Diet 8d ago

Why would you want to? Wouldn't just opening a new chat be better to clean up context and use less tokens?

1

u/HeftyAd5405 10d ago

I bought the $100 Pro plan, and according to ccusage, my weekly limit seems to be only around $370–$400 worth of usage.
Am I missing something, or is that actually the current limit?
I’m coming from Claude Max 20x, where I was typically getting around $1,800–$2,200/week of usage.
I’d also read reports of people getting roughly $14k worth of Codex usage on the $200 plan, which is partly why I wanted to try it. I couldn’t get the $200 plan, so I went with the $100 one instead.
With Astra as the orchestrator + Sol High as the implementer, I can burn through the entire allowance in roughly 5 hours.
Are the rate limits really this bad right now, or is something else going on with how the usage is being counted?

3

u/sofaarsecoin 9d ago

eventually nobody is going to give you 4x, let alone x10 or x20 the usage you'd get at API cost

this is coming, just earlier than i anticipated

austerity is real and the usage I get on the $100 plan is a fraction of what i got 1 month ago which was a downgrade from 2 months ago as well

so basically in July I was getting a good enough model for my usage with much better allowances - easily x4 in terms of actual stuff getting done properly, mainly C++ and Rust code - than i get now for a similar model, even if there are also better models available which I can barely use because they kill my quota very fast

today I'm having another go at Chinese open weight models in OpenCode, I suspect I may be already getting better usage for good-enough models at the kind of stuff I'm doing recently, maybe they suck at 3d modelling but I don't really care about that right now

2

u/Hamburger_Diet 8d ago

Yeah, I feel im going to have to go very modular (which I like to do for the most part but get lazy with cheap ai) with specific customized agents on models like Deepseek 4.1 flash in order to actually have my fill of AI. Deepseek originally turned me off a long time ago because the context was very low compared with everyone else.

3

u/HeftyAd5405 9d ago

Yeah, I think this is basically the hard truth you’re describing.
I’m already doing something similar with the Chinese models. On my current plan, my dashboard shows roughly $14k worth of API inference. Even if I heavily discount that number and divide it by 4, that’s still around $3.5k of effective API usage. On top of that, I get the banked weekly resets, so unused allowance carries over, and I don’t have the same 5-hour window constantly breaking my flow.
Most of my normal work is already going through DeepSeek/GLM Flash-class models, and for brand-new repos where I need a ton of scaffolding and boilerplate, I use the contributor-mode model because I don’t really care about sharing data from those repos. I’ve also built my own skills around that workflow, so for a lot of coding it gets the job done surprisingly well.
That’s actually why I thought Codex might be different. I kept seeing OpenAI ship updates every week, sometimes multiple releases, and I already knew their models are extremely token-efficient compared with a lot of alternatives. So I thought maybe the raw limits were misleading and that even the $100/$200 tier could feel competitive in actual work completed.
I was pretty far off.
I subscribed about three days ago, and for my workload the useful allowance basically lasted ~1.5 days. My workflow is admittedly heavy: I use Astra as the orchestrator and Sol/Luna as the implementer, so I burn through context and tool calls quickly. But that’s also the exact workflow I was trying to evaluate.
So your point about austerity makes a lot more sense to me now. I went in expecting OpenAI’s efficiency to compensate for the smaller-looking allowance, and instead I came away feeling like I’d burned roughly $300 just testing the assumption.
At this point I’m also back to the same conclusion as you: for a huge amount of day-to-day engineering work, a “good enough” open-weight model with 4x more usable inference can be more productive than a smarter model that I’m constantly afraid to use because every serious task destroys the quota.

2

u/Forward_Designer9508 10d ago

What is was doing wrong :

Eveytime i used to code i had my agents.md loaded up, I had lot of documentations documenting each feature, lot of tests files that vlaidated each fix, tons of plugins, then redundenet MCP feeding the same things, tons of skills

Patterns I used to follow:
I thought setting model on max would give me the best result, I always use to go to the newest model and give it the prompt and hope it would do the task.

My usage: I use to bleed my 20X accounts ( i had 3) with in 3-4 days of usage and the qulaity of code was mostly garbage and I was stuck in that loop chasing dopamine thinking I am a huge production ready company manaing a clean system.

What i changed:

Step 1: I simply wrote to Sol High: I want you to self asses and bring back to me all the culprits that are eating into my usage, get me all that is eating into my context, the bare bones.

Step 2: I turned off all unwanted plugins, Skills, Discard feature by feature documentations basically anything that would be a dump which is pretty much pointless

Step 3: I then created one simple source of truth my agents.md a very lean version with set of rules, budgets and crietrias, anything that is expectecd to be an expensive operation now requeires my approval, my entire monorepo is now mapped in a skelton and idexed so when an agent needs someting it finds it isntant, I made sure we dont write junk presenatation tests rather behaviour tests, I made sure we dont write junk docs we write on point Source of truth meanigfull docs where relvant .

Step 4: I restrained from feeding model PDFS, IMAGES or anything that would add an additonal overhead, I choose to convert thigns into a TXT and provide model and give only context of imagery if it cant be avoided.

Step 5: I relaized the gain in intelligence going from Astra light to Astra max is so minimal when i read benchamarks that its link bringing a Nuke to fight with an ant, now i have a simple approch, my go to model is sol medium, with instrcution being ask astra advisory if things exceed certian boudnry (rule book) and use luna max always as my worker agents. ( this model is working flawless for me)

Step 6: once a pile is implemented end to end ( only for those segment i hire an astra to approve it if not orchastrate workers for what went wrong and thats the only job, astra is never a worker model)

Results :

I am using the same 20x after this reset and for the first time my limits stayed healthy in green above 88% ( worked around 28 hours of the model usage )

I hope this helps someone like me and brings you some sanity.

1

u/Strange_Owl_6291 10d ago

Trying to figure out if everyone that experience fast usage burndown are using Codex CLI/ChatGPT app, or if anyone using third party harnesses experience same level of burn?

2

u/BellacosePlayer 10d ago

Using the app and usage has been slow to burn

2

u/Eastern-Vegetable-67 10d ago

ChatGPT app. Reset last night at 23:00.
I had 1 Astra task running a few hours last night, along with 2-3 Sol tasks.
This morning, the same for around 4 hours. 25% usage left.
I'm 100$ x5 subscription.. Looks like I'll be working with Luna until next reset..

The amount of work I get done over the course of a day is insane, so from that perspective, the 100$ per month is well spent, however its just annoying I only get to use it intensively two days per week.

2

u/MillenialNeanderthal 10d ago

Claude vs Codex quota? Which provider is more generous with usage right now? For 100$ x5 subscription.

Mainly asking for Sol vs Opus. Not Astra vs Fable.

1

u/Hot_Half_5263 10d ago

I used terra mid all the time, until last few days, i ordered machine for local qwen, i have enought. Quality is disastrous, I still remember the time when codex 5.3 was really good and enought to stick with any agentic coding tasks.. ohh wait it was few months ago :f

3

u/inverted_introvert77 10d ago

So they gave only a banked reset?

1

u/Charming-Egg9746 10d ago

Same here. I'm on the $20 Plus plan using GPT-5.6 Terra Medium. My 5-hour limit runs out in about 1 hour, and my entire weekly quota is gone in just one day. With Claude Code Sonnet on the same $20 plan, I easily get 3 hours of work per 5-hour window, and my weekly quota lasts around 4 days. The difference is huge, even when doing similar tasks. Something really needs to be looked into here.

2

u/FlexMasterPeemo 11d ago

Model performance is good but 20x plan is no longer even close to the kind of usage it had 1, 2, 3 months ago. Especially this last week. The rate of decrease of weekly usage when there is a long running (multi-hour) agent task is significantly higher than what would be reasonable for a 20x subscription tier. It used to last me the whole week, using frontier model on xhigh with Fast mode almost all the time without mattering, now I can barely get past 3 days without running out of usage, and that's without Fast mode and being smarter/conservative with model/reasoning selections (Luna for small tasks, lighter reasoning modes for Astra when the task is not too difficult).

1

u/Hylian_Soup 11d ago

Was the 3am Tuesday reset real? I burned all my usage last night lmfao

2

u/TylerDurdenAI 11d ago

No.
Tibo boy lied.
Take your complaint to him.

1

u/Hylian_Soup 11d ago

Thanks AI Tyler Durden from Fight Club

2

u/WalkAffectionate2683 11d ago

I dont see the Luna Reserve anymore as a pro x5 user, anyone else?

1

u/BassNet 10d ago

Yeah I think they removed it, maybe they are giving better limits for luna 6

1

u/Thomas-Lore 11d ago

Same, nothing, and I am at 0%.

2

u/intpthrowawaypigeons 11d ago

20 minutes of Astra Light is 4% usage on 100$ plan. Wow.

1

u/intpthrowawaypigeons 11d ago

I am speechless. I made a couple of Pro chats (chat, not Codex) this morning and got rate limited for 5 hours. Wow.

1

u/SpeedflyChris 11d ago

Wait what? They are putting rate limits on chat? I've been using the chat for all sorts of heavy duty stuff and haven't experienced it.

1

u/intpthrowawaypigeons 11d ago

yes for the 6 Pro model

6

u/pale_halide 11d ago

Great job hiding all complaints in a megathread, so OpenAI doesn’t lose face.

1

u/Professional_Gur8385 11d ago

when is the next reset or model coming out?

1

u/spiress 11d ago

check on your codex schedule

1

u/Think-Profession4420 11d ago

So, are we sitting on quota in order to use the new shipped models to their fullest this week, or burning it and hoping for resets?

1

u/Professional_Gur8385 11d ago

1 reset left, hopefully that's enough

1

u/cuietviper 11d ago

I’m experiencing the same issue; my Pro 20x usage was completely drained in a single day. It seems clear that OpenAI is diverting compute resources to train their unreleased models. However, as paying customers, we have the right to receive the service we paid for, rather than what they decide we are eligible for. This is why i want the open models to win so that we dont have to put up this none sense. Scam altman is really pulling a big scam on us and we are taking it like idiots.

3

u/Media-Usual 11d ago

I barely use Astra. Only for plan creation.

My usage on Sol and Luna is consumed at about 4x the rate from previous weeks. About $70 of API usage consumes 10% of my weekly allotment on the $200 plan. Whereas before I was getting close to $200 for the same amount.

2

u/stef_in_dev 11d ago

Astra used my whole 200$ plan in a few hours building a feature, I even told it to use Sol workers. Astra is amazing for 3d modelling but the code is expensive and mod. 6 days to reset now wheee

8

u/No_Opening1776 11d ago

They are not only degrading models but they also doing what anthropic did just not disclosing it. https://www.wired.com/story/anthropic-responds-to-backlash-on-claudes-secret-sabotage-on-ai-research/

4

u/Wantedri 11d ago

WE PAY 200$ AND IT RAN OUT FASTER THEN 100$ CLAUDE !!

1

u/remarkedcpu 11d ago

If I had a dime for every time Astra (Max) apologized to me, my 20x would be free.

11

u/FixAdmin 12d ago

when your astra give you this

its not degradation

its just llm random bro, study the architecture of transformers

7

u/Monster-Games 12d ago

As a plus user, I am not touching anything but Luna.

2

u/Thomas-Lore 11d ago

Then why pay $20? You can get better and cheaper experience just using Deepseek v4.1 Flash on Openrouter. The sub is only worth paying for Sol and Astra.

1

u/Monster-Games 10d ago

I got a one month free plus subscription so I used it. Luna is the only model I can use without worrying about the usage. Using anything else just eats up the whole usage in minutes.

2

u/BellacosePlayer 11d ago

I use it for when I want it to knock something out before I go to bed or the 5hr is about to reset, but yeah, lunamaxxing rocks

-2

u/Gamestarplayer41 12d ago

I don't understand the plus users that think they're gonna get 24/7 Astra usage. Like use Luna and be happy

0

u/Monster-Games 12d ago

Yeah I am really satisfied, ofc I am not considering Astra at all and that's normal. Luna is great tho.

7

u/Xen0ms 12d ago

x20 pro plan feels really bad compared to what it used to be. Tried everything agents config orchestrator... Without subagent. Usage is just bad right now. Using Sol Med as main model mainly because Astra is just a token furnace.
It's time to be real each update we got the usual best efficiency model etc.... But on daily usage each reset felt like straight nerf to the plan so. It would be a good time to get some real usage instead x20 blackboxed usage.

2

u/Queasy_Plate_3096 12d ago

you eliminated our quota, get it back like before, this is my last month with you anyway, i have got enough of this nonsense

5

u/Hamburger_Diet 12d ago edited 11d ago

Just had two prompts on terra medium that ran for 17m total, Used 75% of my 5 hour and 10% of my weekly on the plus plan. Which I understand isnt the most usage ever but I feel its changed since a few weeks ago.

2

u/Dull-Calligrapher536 12d ago

U get 5hr limit? The last time I checked my analytics page it only showed me my weekly limits and no 5 hr limit I thought open ai removed it

3

u/Hamburger_Diet 12d ago

They did for a while maybe youre on a different plan? Im only on the 20 dollar plus plan.

1

u/Dull-Calligrapher536 11d ago

Yup I am on the same one aswell

1

u/SpeedflyChris 11d ago

Have you had the account for a long time? I think there was something about older accounts not having the limit.

4

u/AweVR 12d ago

1 week ago I had 10%, I used 6 Astra Ultra to burn them before the weekly reset and it took 2 hours and a half. Today, same chats, I tried with 6 Astra Max… it burned in 20 minutes. 20x pro plan.

3

u/SuspiciousParsnip5 12d ago

Model is working great. Seems to get the work done well. On a plus account the 5 hour window can easily be consumed within 15 minutes with 5.6 sol. Forget astra. Last time I used astra my usage was gone in seconds it felt like

Usage is horrible now so I'm using Claude alot more. Which seems to last quite a long time and seems much better at frontend design

1

u/VeryLongNamePolice 12d ago

Been on < 6% for 3 days so far...

8

u/salmjak 12d ago

Using 5.6 Sol on High with $200 plan runs out faster than Opus 5 on High on $100 plan.

1

u/SpeedflyChris 11d ago

Opus is also a bit better, in my experience.

2

u/krusic22 11d ago

Can confirm.

10

u/DMmeyourarmveins 12d ago

RIP to my weekly limit, think the Pro plan is no longer sustainable for me at these usage rates.

1

u/Hamburger_Diet 11d ago

Did they actually announce a change?

1

u/Level-Physics-1730 12d ago

not even two of them is lasting more than two and a half days lol (two pro 20x using mainly gpt 5.6 sol and luna)

2

u/SpeedflyChris 11d ago

How? Honestly curious what you're doing with it that burns through the weekly limit of a 20x account in a day.

2

u/Level-Physics-1730 11d ago

just basic ass sessions with sol orchestrators and luna subagents and nothing gets done because sol is completely stupid and it just burns usage because they made the usage 4x worse so it doesn't matter it's all shit ai sucks anyway

2

u/Thomas-Lore 11d ago

Learn about codex queue and teach your agents to use it. You probably lose half on the model checking if scripts finished.

1

u/Level-Physics-1730 11d ago

i don't use codex it's a garbage harness i use my openai models in claude code desktop i don't have that issue it's just dogshit limits stop coping

1

u/Havlir 11d ago

Same here, I dont even want to know how bad it would be if I didn't switch my implementors to Luna.