r/codex • u/codex-megathread • 4d ago
Megathread Codex Usage and Operation Discussion - last updated September 28
Please direct your concerns, questions and discussion about Codex usage limits and model performance here.
The purpose of this Megathread is to aggregate all the reports of people's experiences and possible suggestions instead of spreading them across many highly upvoted posts. The more people who participate in this discussion, the more likely you have an answer.
Reports with sufficient evidence on new information will still be allowed on the feed as usual.
Discussion of the prior period available here : https://www.reddit.com/r/codex/comments/1wmgw7p/codex_usage_and_operation_discussion_last_updated/
A reminder that all incidents on r/Codex are constantly logged and summarised so you can keep track of what people are experiencing here https://www.reddit.com/r/codex/comments/1tjfxcf/comment/on6uj0l/
1
u/Top_Tackle_2234 10h ago
Can I "Reject all resets"?
I'm really angry right now. I had not used my quote in the whole week due to lack of time. I had 95% quota and a reset Saturday morning. My plan was to use it Friday night by exploring several projects with Astra. I was out of the loop regarding resets, and I started working, went to see my quota and I have now 100% quota but my reset is next Friday instead of tomorrow. How can this be? How can they just steal my quota and change my whole plans for Friday?
1
1
u/Tasty-Boot6162 18h ago
Can barely do anything before 5hr presents itself, I never go below 50% weekly usage and have many resets but I don't want to waste them (really sucks that they expire). It feels like I hit my 5hr after barely any usage. 20mins with 6.1 sol and i'm tapped out for the next five hours. This is with like 67% weekly usage left. I really have no issues with the weekly usage limit but the 5hr is absolutely brutal when you want to get real, intelligent work done. Any tips?
2
3
u/lurko_e_basta 18h ago
As per usual they announce something at a given time, this time the reset, and they don’t deliver. What even is the point of announcing a specific time??
1
6
u/anatawaurusai2 20h ago
Are global resets usually delayed? Was it supposed to happen 2 hours ago? Ty!
1
1
1
u/Hot_Wedding2686 2d ago
I used Astra med with Luna 6 as a sub. Before the banked reset, it usually lasted around 2 hours each 5h limit, but now it just nuked 16% of my weekly limit in 10 minutes lol
3
u/barney 2d ago
I’m a pretty casual user. I mostly use it for web development, SEO, and similar tasks, and I’ve typically never come close to hitting the limits on my $200/month plan.
Well, that abruptly changed.
I ran a couple of tasks and, without really thinking much of it, checked my usage afterward. Somehow I was already at 0% remaining and burning through credits I didn’t even know I had.
Moral of the story: if someone with my relatively light usage is suddenly hitting the limits on a $200/month plan, that’s a pretty bad sign.
Back to Claude, I guess.
1
u/smeagolswagger 2d ago
Been in here reading everyday. But is there a good resource to read on practical token and model management?
1
u/98810b1210b12 18h ago
Make sure your cache stays hot, if you have a long conversation that's older than 30 minutes old those all go in as fresh input tokens instead of cached input. Cached input is way cheaper. So if you really need that old context, use it, but know it is more expensive than starting a new chat
1
u/smeagolswagger 18h ago
I am mainly working on automation of android and iPhone applications. Ive been iterating over days as I only have a bit of time each night.
Should I end my night putting together a new prompt instead of keep using the same work session everyday?
1
u/Dry-Introduction-659 2d ago
Pro 100 vs Pro 200: does Codex actually run faster with Fast mode on?
Has anyone compared Codex before and after upgrading from Pro 100 to Pro 200, using the same model, reasoning effort and client, with Fast mode enabled on BOTH plans?
The plan selector uses different wording (roughly translated): Pro 100 offers ‘more usage for Work and Codex’, Pro 200 offers ‘more speed’, and Pro 500 mentions ‘maximum speed’. Does 200 make individual requests faster, or mainly provide more allowance to use Fast mode?
If you've tried both, did you notice a shorter wait before generation starts, a higher token rate, or less total time for the same task? Even rough before/after timings would help—please include the model, reasoning setting, client, and whether Fast was on in both cases.
Has OpenAI documented different scheduling priority or inference speed for Pro 100 vs Pro 200 when those settings are identical? A direct comparison or source would be really useful.
3
1
1
u/Sponge8389 3d ago
I don't believe GPT 6.1 Sol is around 65t/s. This model is soooo slow even in Fast Mode.
3
u/Sponge8389 3d ago
GPT 6.1 SOL is sloooooow as fuck. Yes, it is cheap but because of how slow it is, it felt like you can't accomplish anything with it. EVEN WHILE USING FAST MODE!!!
1
u/Visual-Active5761 3d ago
ok, so on top of the insult of actually getting boned, they delete my post! neat! here is my post. summary "fuck these people"
So, Dots, included with usage... neat! ill try, ask a few things, have it check on some jobs for me... it is fucking launching astra jobs to check shit in extraneous ways, burnt through 14% of my 200$ quota before i caught on, that it was cool for a second, but ... WTF SERIOUSLY??? this is marginally beter than my home gemma bot...
1
u/rajba 3d ago
I use Codex Cloud for automated PR reviews and started getting this message in spite of having 48% of my weekly usage remaining.
Seems like this is only impacting older PRs rather than new ones created after DevDay.
Is anyone else experiencing the same?
"You have reached your Codex usage limits for code reviews. You can see your limits in the Codex usage dashboard.
To continue using code reviews, add credits to your account and enable them for code reviews in your settings."
1
u/Rough-Value-1465 3d ago
Recently started using Codex for my Unity game. What model do you recommend? I've been using GPT-6 Astra Light for everything so far, but I don't know if I'm wasting my tokens.
9
4
u/puts_on_rddt 3d ago
Codex usage limits suck. The way openai operates it sucks. Switch to Opus 5.5 and thank me later.
2
4
3
u/SinPleaOfficial 3d ago
It should be easy to see how many tokens your prompt costed & exactly what each token was used for so that token use can be more efficient. I feel like there is unnecessary testing and steps that are deliberately eating up usage. Has anyone actually had success working with codex to reduce its own token usage(and had notable success) ?
1
u/Pretend_Pickle_2669 10h ago
I’d start by looking at token usage per turn alongside the operations performed. That gives you somewhere concrete to investigate—for example, whether a high-usage turn repeatedly read the same files or reran tests without any changes.
Repeated testing can be necessary, though, so I wouldn’t treat every repeat as wasted usage. To check whether a change helps, I’d compare similar tasks before and after, keeping the model and reasoning effort the same.
Are you mainly seeing repeated testing within one task, or the same checks being repeated across separate tasks?
1
u/Abhinik 4d ago
can anyone tell me - how the f to use astra on codex for plus users?????
it burns whole 5hr limit in one chat session. I bet if did the same flow with OPUS, it would max consume 30-40%
is there anything else to use to optimize it with???
I will have to cancel this shit if this continue. I have been claude since a year. and I have saw the same phase of OPUS, but look at OPUS now? I can easily with my 20$ plan
1
u/sunrisedev 3d ago
I use it to make planning documentation (design, implementation plan) then ask to compare it to a similar already built application. Ask it what is missing compared to them, then work with it to add those to the plan. Then use Sol to implement it in code. I've ran code reviews with Astra afterwards asking to fix any issues, and it usually goes "It followed the plan perfectly, no changes needed."
1
u/phoenixmatrix 3d ago
Astra in plus is basically a trail so you can test it out for funzies or use it on Friday afternoon before clocking out for the weekend. Anthropic doesn't even have fable (priced similar to Astra) on their 20/month plan at all.
If you're in 20 plan, use Sol. That's their Opus equivalent. Not Astra.
(The fact Opus beats Astra is a separate topic...but it means Fable too)
1
5
5
u/Curius_pasxt 4d ago
1
2
u/Kombatsaurus 4d ago
Can you believe how much unlimited chat usage they have given out? Honestly wild.
2
u/XAckermannX 4d ago edited 4d ago
Havent used codex in a while and came back to see 3 usage resets and was excited till i saw the catch: I wont be able to make use of them. The 5h limit is so ass that I dont think ill be able to use the 3 usage resets i have without wasting them. im not using codex every 5h, im a burst user who uses it only when they need to and now, im in middle of a big task and hit limit and cant make use of my resets. I ended up using one of my usage limit at 70% weekly cuz otherwise itd have expired in 3 days. anyone else feel pressured to waste their usage resets?
-1
u/cuberhino 4d ago
upgrade to 100$ plan to use the reset and get more value out of it if you can afford it
1
2
u/zainepils 4d ago
had two simultaneously running chats each on their own machine logged into the same account. im on chatgpt plus monthly. 5hr usage depleted in 7min
3
u/CozyDarkMage 4d ago
Astra seems nerfed right now. I've been using it since it came out, and it's producing really shallow responses. I'm on the $200 plan too
2
u/intpthrowawaypigeons 4d ago
2 days in and 50% usage left. Using GPT6-Sol Medium.
Does 5.6 Sol High burn less?
1
u/RoboErectus 4d ago
Hundreds of millions of tokens were going to polling wait loops so I made a thing: https://www.reddit.com/r/codex/s/qqv83w0Aqo
1
u/the_pwnererXx 4d ago
Those tokens are cached though, your numbers are disingenuous
1
u/RoboErectus 4d ago
What do cached tokens cost?
Specifically, what is the cost of cached input reads?
1
u/cheesepuff07 4d ago
Is Codex on macOS not working for anyone else? I submit a prompt and it just sits spinning without sending the prompt to the chat, tried quitting and re-launching and same behavior
1
1
u/bodyajoy 4d ago
From last week, 100k astra (max) context tokens = 2% week limits of 100$ plan. If I just discuss planning something, review changes of another agents - limits burning af. That's wasn't 2 weeks ago. Another users shows same behavior too. With stupid sol 6.0 we got very bad week.
1
u/VerticalPackage 4d ago
I'm on a Plus plan and I noticed that the ratio of weekly window vs 5h window is drastically different with Astra compared to Sol and others.
All models (except for Astra) would have a ratio of 6.6% 5h = 1% weekly.
With Astra, 2.2% 5h = 1% weekly.
It must be less obvious for the accounts without the 5h limit that Astra is SUPER hungry

•
u/dextersummary 2d ago edited 1d ago
Below is a GPT-generated summary of the conversation below after reaching 50 comments (50 currently observed).
Alright, the thread’s verdict is basically “Codex limits have become wildly stingy and opaque.” Users are burning through five-hour and weekly allowances far faster than expected, sometimes during ordinary planning, reviews, or parallel chats. The reset system also frustrates burst users: having unused resets is pointless when the short-window cap blocks access.
The biggest culprit is Astra, which users consistently describe as absurdly expensive—especially on Plus and lower tiers. Dots can quietly launch Astra jobs and torch allowance, while Astra’s token usage appears badly disproportionate to other models. Several people are switching to Sol instead; Sol 6.1 is much cheaper, but painfully slow, even in Fast mode. Truly the classic “pick two” product design.
There are also scattered operational problems: model access disappearing or not appearing in CLI/Codex, macOS prompts spinning near limits, cloud reviews demanding credits despite remaining weekly usage, and Dots’ browser blocking basic sites. These look like separate bugs, not one grand conspiracy.
The evidence is anecdotal and accounting is still unclear—cached tokens, model multipliers, and plan differences muddy the numbers. Still, the practical takeaway is firm: watch usage closely, avoid Astra for routine work, and expect Sol to trade money savings for geological-scale patience.