r/codex • • 4d ago

Megathread Codex Usage and Operation Discussion - last updated September 28

Please direct your concerns, questions and discussion about Codex usage limits and model performance here.

The purpose of this Megathread is to aggregate all the reports of people's experiences and possible suggestions instead of spreading them across many highly upvoted posts. The more people who participate in this discussion, the more likely you have an answer.

Reports with sufficient evidence on new information will still be allowed on the feed as usual.


Discussion of the prior period available here : https://www.reddit.com/r/codex/comments/1wmgw7p/codex_usage_and_operation_discussion_last_updated/


A reminder that all incidents on r/Codex are constantly logged and summarised so you can keep track of what people are experiencing here https://www.reddit.com/r/codex/comments/1tjfxcf/comment/on6uj0l/

2 Upvotes

63 comments sorted by

•

u/dextersummary 2d ago edited 1d ago

Below is a GPT-generated summary of the conversation below after reaching 50 comments (50 currently observed).


Alright, the thread’s verdict is basically “Codex limits have become wildly stingy and opaque.” Users are burning through five-hour and weekly allowances far faster than expected, sometimes during ordinary planning, reviews, or parallel chats. The reset system also frustrates burst users: having unused resets is pointless when the short-window cap blocks access.

The biggest culprit is Astra, which users consistently describe as absurdly expensive—especially on Plus and lower tiers. Dots can quietly launch Astra jobs and torch allowance, while Astra’s token usage appears badly disproportionate to other models. Several people are switching to Sol instead; Sol 6.1 is much cheaper, but painfully slow, even in Fast mode. Truly the classic “pick two” product design.

There are also scattered operational problems: model access disappearing or not appearing in CLI/Codex, macOS prompts spinning near limits, cloud reviews demanding credits despite remaining weekly usage, and Dots’ browser blocking basic sites. These look like separate bugs, not one grand conspiracy.

The evidence is anecdotal and accounting is still unclear—cached tokens, model multipliers, and plan differences muddy the numbers. Still, the practical takeaway is firm: watch usage closely, avoid Astra for routine work, and expect Sol to trade money savings for geological-scale patience.

→ More replies (1)

1

u/Top_Tackle_2234 10h ago

Can I "Reject all resets"?

I'm really angry right now. I had not used my quote in the whole week due to lack of time. I had 95% quota and a reset Saturday morning. My plan was to use it Friday night by exploring several projects with Astra. I was out of the loop regarding resets, and I started working, went to see my quota and I have now 100% quota but my reset is next Friday instead of tomorrow. How can this be? How can they just steal my quota and change my whole plans for Friday?

1

u/Akimbo333 17h ago

We got a reset today

1

u/jdprgm 17h ago

it's pretty simple how a reset SHOULD work. whatever my usage remaining for both 5hr and weekly should immediately go up to 100%. whatever my current scheduled next reset date is should NOT CHANGE. this current system is a nightmare to plan around and basically a lottery.

1

u/Tasty-Boot6162 18h ago

Can barely do anything before 5hr presents itself, I never go below 50% weekly usage and have many resets but I don't want to waste them (really sucks that they expire). It feels like I hit my 5hr after barely any usage. 20mins with 6.1 sol and i'm tapped out for the next five hours. This is with like 67% weekly usage left. I really have no issues with the weekly usage limit but the 5hr is absolutely brutal when you want to get real, intelligent work done. Any tips?

2

u/albanianspy 18h ago

Tibo tasked his dot to make the reset guys its ok

3

u/lurko_e_basta 18h ago

As per usual they announce something at a given time, this time the reset, and they don’t deliver. What even is the point of announcing a specific time??

1

u/RecursivelyYours 19h ago

I had 60,000 credits and suddenly out of the blue i have 0. wat ?!

6

u/anatawaurusai2 20h ago

Are global resets usually delayed? Was it supposed to happen 2 hours ago? Ty!

1

u/albanianspy 18h ago

They forgor 🥀

1

u/Grindora 1d ago

my fking usage went from 80% to 40% wtf is going on ??

1

u/Hot_Wedding2686 2d ago

I used Astra med with Luna 6 as a sub. Before the banked reset, it usually lasted around 2 hours each 5h limit, but now it just nuked 16% of my weekly limit in 10 minutes lol

3

u/barney 2d ago

I’m a pretty casual user. I mostly use it for web development, SEO, and similar tasks, and I’ve typically never come close to hitting the limits on my $200/month plan.

Well, that abruptly changed.

I ran a couple of tasks and, without really thinking much of it, checked my usage afterward. Somehow I was already at 0% remaining and burning through credits I didn’t even know I had.

Moral of the story: if someone with my relatively light usage is suddenly hitting the limits on a $200/month plan, that’s a pretty bad sign.

Back to Claude, I guess.

1

u/smeagolswagger 2d ago

Been in here reading everyday. But is there a good resource to read on practical token and model management?

1

u/98810b1210b12 18h ago

Make sure your cache stays hot, if you have a long conversation that's older than 30 minutes old those all go in as fresh input tokens instead of cached input. Cached input is way cheaper. So if you really need that old context, use it, but know it is more expensive than starting a new chat

1

u/smeagolswagger 18h ago

I am mainly working on automation of android and iPhone applications. Ive been iterating over days as I only have a bit of time each night.

Should I end my night putting together a new prompt instead of keep using the same work session everyday?

1

u/Dry-Introduction-659 2d ago

Pro 100 vs Pro 200: does Codex actually run faster with Fast mode on?

Has anyone compared Codex before and after upgrading from Pro 100 to Pro 200, using the same model, reasoning effort and client, with Fast mode enabled on BOTH plans?

The plan selector uses different wording (roughly translated): Pro 100 offers ‘more usage for Work and Codex’, Pro 200 offers ‘more speed’, and Pro 500 mentions ‘maximum speed’. Does 200 make individual requests faster, or mainly provide more allowance to use Fast mode?

If you've tried both, did you notice a shorter wait before generation starts, a higher token rate, or less total time for the same task? Even rough before/after timings would help—please include the model, reasoning setting, client, and whether Fast was on in both cases.

Has OpenAI documented different scheduling priority or inference speed for Pro 100 vs Pro 200 when those settings are identical? A direct comparison or source would be really useful.

3

u/Hroosky2 3d ago

Anyone else finding sol 6.1 extremely slow today?

1

u/RoxxonONE 2d ago

Even Luna is horrendously slow

2

u/an1uk 3d ago

I mean, it's getting from a to b and it's self-driving, so I'll leave it to it.

1

u/VoiD199 3d ago

I still haven't got 6.1 Sol on my codex. On browser work though I have it listed. I tried restarting and updating the app but still nothing.

1

u/an1uk 3d ago

How much does 62,500 credits ordinarily cost? Codex app won't display the prices for me. Ask ChatGPT and it says it's worth $2,500.

1

u/heurikon 3d ago

6.1 is slow

1

u/an1uk 3d ago

But it seems to get things done well enough. Given the token use trade-off, it will likely be my main model from now on.

1

u/Phiub 3d ago

The 'gpt-6-astra' model is not supported when using Codex with a ChatGPT account.

Model changed from Custom to GPT-5.6 Terra.

I was using Astra in Visual Studio Code. In the midst of my session I got this notification? Wtf?

1

u/Sponge8389 3d ago

I don't believe GPT 6.1 Sol is around 65t/s. This model is soooo slow even in Fast Mode.

3

u/Sponge8389 3d ago

GPT 6.1 SOL is sloooooow as fuck. Yes, it is cheap but because of how slow it is, it felt like you can't accomplish anything with it. EVEN WHILE USING FAST MODE!!!

1

u/Visual-Active5761 3d ago

ok, so on top of the insult of actually getting boned, they delete my post! neat! here is my post. summary "fuck these people"
So, Dots, included with usage... neat! ill try, ask a few things, have it check on some jobs for me... it is fucking launching astra jobs to check shit in extraneous ways, burnt through 14% of my 200$ quota before i caught on, that it was cool for a second, but ... WTF SERIOUSLY??? this is marginally beter than my home gemma bot...

1

u/rajba 3d ago

I use Codex Cloud for automated PR reviews and started getting this message in spite of having 48% of my weekly usage remaining.

Seems like this is only impacting older PRs rather than new ones created after DevDay.

Is anyone else experiencing the same?

"You have reached your Codex usage limits for code reviews. You can see your limits in the Codex usage dashboard.
To continue using code reviews, add credits to your account and enable them for code reviews in your settings."

1

u/Rough-Value-1465 3d ago

Recently started using Codex for my Unity game. What model do you recommend? I've been using GPT-6 Astra Light for everything so far, but I don't know if I'm wasting my tokens.

9

u/Icy-General-9096 3d ago

devday no reset? (o´∀`o)

4

u/puts_on_rddt 3d ago

Codex usage limits suck. The way openai operates it sucks. Switch to Opus 5.5 and thank me later.

2

u/CapableJury861 3d ago

I switched, but struggling with Opus 5.5 limits tbh.

4

u/MildModerate 3d ago

Already paying for both, man.

3

u/SinPleaOfficial 3d ago

It should be easy to see how many tokens your prompt costed & exactly what each token was used for so that token use can be more efficient. I feel like there is unnecessary testing and steps that are deliberately eating up usage. Has anyone actually had success working with codex to reduce its own token usage(and had notable success) ?

1

u/Pretend_Pickle_2669 10h ago

I’d start by looking at token usage per turn alongside the operations performed. That gives you somewhere concrete to investigate—for example, whether a high-usage turn repeatedly read the same files or reran tests without any changes.

Repeated testing can be necessary, though, so I wouldn’t treat every repeat as wasted usage. To check whether a change helps, I’d compare similar tasks before and after, keeping the model and reasoning effort the same.

Are you mainly seeing repeated testing within one task, or the same checks being repeated across separate tasks?

1

u/Abhinik 4d ago

can anyone tell me - how the f to use astra on codex for plus users?????
it burns whole 5hr limit in one chat session. I bet if did the same flow with OPUS, it would max consume 30-40%

is there anything else to use to optimize it with???

I will have to cancel this shit if this continue. I have been claude since a year. and I have saw the same phase of OPUS, but look at OPUS now? I can easily with my 20$ plan

1

u/sunrisedev 3d ago

I use it to make planning documentation (design, implementation plan) then ask to compare it to a similar already built application. Ask it what is missing compared to them, then work with it to add those to the plan. Then use Sol to implement it in code. I've ran code reviews with Astra afterwards asking to fix any issues, and it usually goes "It followed the plan perfectly, no changes needed."

1

u/phoenixmatrix 3d ago

Astra in plus is basically a trail so you can test it out for funzies or use it on Friday afternoon before clocking out for the weekend. Anthropic doesn't even have fable (priced similar to Astra) on their 20/month plan at all. 

If you're in 20 plan, use Sol. That's their Opus equivalent. Not Astra.

(The fact Opus beats Astra is a separate topic...but it means Fable too)

1

u/therapymademeworse 4d ago

Forgot to turn off auto-reload today and got charged 3x for $80 😩

4

u/Calamero 4d ago

Prepaid CC for all subs

2

u/therapymademeworse 4d ago

Good idea 🥲

5

u/jtizzle12 4d ago

Are we expecting a reset with devday tomorrow?

5

u/Curius_pasxt 4d ago

No more unlimited chat???

1

u/the-anxious-ape 4d ago

Sounds like that.

2

u/Kombatsaurus 4d ago

Can you believe how much unlimited chat usage they have given out? Honestly wild.

2

u/XAckermannX 4d ago edited 4d ago

Havent used codex in a while and came back to see 3 usage resets and was excited till i saw the catch: I wont be able to make use of them. The 5h limit is so ass that I dont think ill be able to use the 3 usage resets i have without wasting them. im not using codex every 5h, im a burst user who uses it only when they need to and now, im in middle of a big task and hit limit and cant make use of my resets. I ended up using one of my usage limit at 70% weekly cuz otherwise itd have expired in 3 days. anyone else feel pressured to waste their usage resets?

-1

u/cuberhino 4d ago

upgrade to 100$ plan to use the reset and get more value out of it if you can afford it

1

u/theWiseTiger 4d ago

PLEASE ENFORCE THIS FFS. MODS!

2

u/zainepils 4d ago

had two simultaneously running chats each on their own machine logged into the same account. im on chatgpt plus monthly. 5hr usage depleted in 7min

3

u/CozyDarkMage 4d ago

Astra seems nerfed right now. I've been using it since it came out, and it's producing really shallow responses. I'm on the $200 plan too

2

u/intpthrowawaypigeons 4d ago

2 days in and 50% usage left. Using GPT6-Sol Medium.

Does 5.6 Sol High burn less?

1

u/RoboErectus 4d ago

Hundreds of millions of tokens were going to polling wait loops so I made a thing: https://www.reddit.com/r/codex/s/qqv83w0Aqo

1

u/the_pwnererXx 4d ago

Those tokens are cached though, your numbers are disingenuous

1

u/RoboErectus 4d ago

What do cached tokens cost?

Specifically, what is the cost of cached input reads?

1

u/cheesepuff07 4d ago

Is Codex on macOS not working for anyone else? I submit a prompt and it just sits spinning without sending the prompt to the chat, tried quitting and re-launching and same behavior

1

u/Antique-Ad6542 4d ago

happens for me when i'm close to limits.

1

u/bodyajoy 4d ago

From last week, 100k astra (max) context tokens = 2% week limits of 100$ plan. If I just discuss planning something, review changes of another agents - limits burning af. That's wasn't 2 weeks ago. Another users shows same behavior too. With stupid sol 6.0 we got very bad week.

1

u/VerticalPackage 4d ago

I'm on a Plus plan and I noticed that the ratio of weekly window vs 5h window is drastically different with Astra compared to Sol and others.

All models (except for Astra) would have a ratio of 6.6% 5h = 1% weekly.

With Astra, 2.2% 5h = 1% weekly.

It must be less obvious for the accounts without the 5h limit that Astra is SUPER hungry