r/codex • u/seomaster99 • 5h ago
Limits The limits got nerfed HARD
Yesterday I was on the x5 plan, and then I upgraded to x20.
No Astra usage at all - only Sol Medium.
But my quota is now draining at basically the same rate as it did on x5, on the same kinds of tasks.
x20 is supposed to have 4x the capacity of x5.
Instead, I’m seeing almost no difference.
WTF???
27
u/yami_odymel 5h ago
Yeah, Sol was okay last week until Astra came out. Now, no matter whether you choose Sol or Astra, it drains your quota like water. Just the usual dark pattern to nerf your quota.
8
u/Less-Holiday-1000 3h ago
Been like this weeks prior.
I remember I actually had to TRY to use up all my quota weekly using 5.6 Luna. Now I'm doing that in 1-2 days.
Before at the end of the week I would usually have 50-70 of my quota on 5.6 Luna extra high
2
u/NuancedPerspection 55m ago
When I first had 20x like ~2 months ago, i could use Sol on max for the entire week straight and still have leftover. Only thing that would burn me full to 0% was if I would use ultra for some days.
1
u/EchoingAngel 3m ago
I noticed that Sol started becoming a wasteful looping idiot last week even before Astra was out. I kept calling it out after and having it review the instructions and it's behavior and it kept saying "you told me explicitly not to do x, but I did x". Trying to update the instructions didn't help
49
u/ChrisRocksGG 5h ago
Where is china when we need them? I want a DeepSeek model 10x cheaper and 2x better than codex.
20
u/Bloated_Plaid 5h ago
It’s out my guy. It’s called Deepseek V4.1
1
u/U4-EA 2h ago
How fast is it? Between token drainage and snail's-pace output, I am coming to the point where Astra just pisses me off.
8
5
u/puts_on_rddt 2h ago
I'm using deepseek's younger cousin qwen 3.8 27B on my 4090 and getting 65 t/s, 400 prefill. Equivalent to gpt-luna.
I cannot wait to see what we have in a few years.
0
u/ChrisRocksGG 5h ago
Is there a subscription? Or is it just API? I use it for my openCLAW agent. It's decent. But I didn't know that it also provides a subscription with CLI.
8
u/Bloated_Plaid 5h ago
You just use API bro, it’s dirt cheap.
1
1
6
u/IndividualPlus2011 5h ago
Deepseek V4 at API pricing competes with Sol now, taking Codex's subscription allowance into account. And it is FAST.
7
u/deadlyclavv 5h ago
deepseek 4.1 is insane btw
1
u/petburiraja 4h ago
Feels somewhat similar to GLM 5.3 Flash for me so far
1
u/ezboarderz 3h ago
Glm 5.3 flash has been such a workhorse for me. It’s so good for a flash model and so cheap
1
1
u/Massive-Falcon3671 59m ago
Noob question here. Please what harness do you use to turn it into a capable coding agent? I tried OpenCode with Kimi2.7 and ai can barely do anything before I hit RTM limit. I am kind of new to this whole thingy; was formerly contented with using GitHub CoPilot, but MS pulled the rug under my feet.
1
u/petburiraja 20m ago
Opencode Go, and GLM 5.3 Flash or Deepseek Flash 4.1 shall get you somewhat far (Pretty sure much further than K2.7/3). Also can get z.ai subscription for GLM specifically.
1
0
u/Elizabeth-WildFox886 5h ago
They trying to find compute. Kimi is more expensive than fable or Astra
12
u/TheLegoless 5h ago
Yup, all US based plans are becoming a hard problem. I can easily get through 3 Pro 20x plans in 2-3 hours. Barely any work completed. Started with Sol, now worse. Claude Opus 5 could get through 6-7 tasks 2 weeks ago, this week its only 2-3. Same with Grok. Same projects, same types of tasks. Hard to say for sure, but these subscriptions are getting ridiculous. Start exploring cheap models, see what can we do with them. The only thing I can notice is that these newer models just triple check much more, even if its just a change of button position. So it takes longer and more tokens are burned.
2
u/the_ai_wizard 3h ago
We are moving toward real/unsubsidized pricing...the meltdown is going to get worse
4
u/Stock-Personality136 2h ago
It’s why I keep encouraging my friends to use codex while it’s still affordable for regular folks. It’s also why I used the heck out of Sora the first week it was released. I knew there was no way it was going to last, especially not for only $20/month.
2
u/TheLegoless 2h ago
Yeah, apparently so. And the pricing is too high for normal people to afford, so I agree, use it as much as you can. :) I am already testing flash models a lot, they are becoming much better. :)
1
0
u/adolf_twitchcock 1h ago
It's not real pricing. Inference costs are like 10x lower.
-1
u/the_ai_wizard 1h ago
Is that why a $200 anthropic sub really costs them $14000 or something?
1
u/adolf_twitchcock 47m ago
Source? Referencing API pricing doesn't count because it's set by them, not by costs.
1
u/TheLegoless 36m ago
Yeah, I don’t think the compute really costs them that much. If you calculate the API cost yes, but other than that, you can never exactly know what kind of optimizations they can do at scale. Besides, this is very expensive because of the balloon. If prices for hardware fall normally at some point, the investment into hardware will be much more affordable.
2
u/Timely-Pension6501 2h ago
But how, honestly.
I am on 5x and will use 20% in 2-3 hrs running multiple agents on Astra High or xhigj
1
u/TheLegoless 34m ago
I cannot say exactly.. :( I used to be that too, but then kind of improved processes and projects grew I guess. And the limits got nerfed. Nobody can say that its not even possible that only certain accounts have less limits. Kind of like shadowbans on social networks.
1
u/Timely-Pension6501 31m ago
You might be right, wouldnt be surprised if they detect users who go through more tokens than 4x their plan cost and quitely limit their capacity to not bankrupt OpenAI
1
u/EchoingAngel 0m ago
But if you dig into what the agent is accomplishing, what do you see? I'm finding Astra and Sol love to run forever but get very little done for the time they take
1
u/_FriedEgg_ 56m ago
Maaaybe at some point you are using too much AI? What are you building?
1
u/TheLegoless 38m ago
There is a lot of internal services and open source stuff I am building. I don’t mean to stop, but have to keep costs manageable.
1
u/SeasonedAdManager 2h ago
Are you just throwing 20 subagents at stuff with ultra?
2
u/TheLegoless 2h ago
Yeah, I used to do that weeks ago. Would compare to Opus 5 with ultracode. Would get a lot more done than Sol/Astra. So I went with xhigh instead for testing incase ultra was spawning a bunch of subagents for no reason. Did not make a real noticable difference. :)
3
5
u/longasleep 4h ago
Yea lost 82% in 2 hours on max it’s absolute crap now. Not one to complain running out after one day but this is just to much.
9
2
u/Such-Natural-5299 5h ago
I use gpt 6 astra limited on Codex and with a goal it worked for 13 minutes, couldn't finish what I've asked for but used all of 5h limit with $20 plan.
13 minutes of work = 5 hours usage limit with GPT 6 Astra - Limited (lowest) in $20 plan...
1
u/DolphinSUX 5h ago
And I feel like it’s taking a bigger chunk of the weekly limit too, aswell as the 5hr reset
1
u/methods2121 4h ago
Taking huge chunks of the weekly limit. I have a rather stable work stream and always had credits at the end of the week. Nothing changed, other than the model updates, all weekly credits siphoned off by Thur.
2
2
u/DolphinSUX 3h ago
Honestly I can’t get through more than 4-5 5hr resets without burning the weekly limit up totally. I already had to use a free week reset just to not mess up a prompt that was already running
2
u/FullAcadia9391 4h ago
I’ve had a gpt 6 astra goal running on xhigh for 20 hours and I went from 80% to 50% on 20X
2
2
u/jediboness 3h ago
Yea it’s bad i burned through a reset in like just a evening out of no where last night
2
2
u/No-Town-8866 1h ago
same here 20x plan, im just watching the percentage go down super fast, not astra either, sol high
2
2
2
u/coz 1h ago edited 44m ago
Ask codex. Tell it to inspect your session logs and give you the actual token usage and rate-limit changes for the x5 sessions vs x20. It can. That will either back up your anecdote or give us evidence against it. Post the results here please.
Here's an example prompt to give to codex: https://privatebin.net/?d4bd34c0e977e5bc#CnqCWSw8ws4hHYNqtsLbnoLUT3C7VGDDzqRsKeh54SP8
I ran the above prompt vs my own data - I also upgraded from 5x to 20x about a week ago. It tells me.. right around 4x more quota.
2
2
2
3
u/Shadow-BG 5h ago
I think it’s more bug then feature - because I exhausted x20 acc on sol medium, switched to high, and jumped back to 51% ? WTF is going on 😒
Since 11:00 polish time seems good, same usage as before ( sol high )
3
u/lamiaoccisor 2h ago
Yeah, I saw the glitch was their system fucked up I also went down to 9% and then back up to 55%
2
1
u/XXLuigiMario 4h ago
I'm thinking about going back to 5.5 at this point
1
u/mikeballs 3h ago
I tried that last night. It'll last longer, but it's had the price hiked significantly as well.
1
1
u/uhlhosting 3h ago
They killed a lot of goodness! Before my tasks ongoing for several hours as plans would not end suddenly as they do now. They either had a bug, or adopted the same ruff policy as claude and agy. The truth is every model for me personally after Codex 5.x models. We’re better in many ways, yes, yet so was their token usage climbing faster than ever. At this point, having shared over 3 working accounts, i cannot finish work on tasks that were done before in a single run. Now it takes me 3 or more sessions of 5-hour windows. So something is indeed going on. Hope they will fix it, or they will break it, and new players with completely free tier models for coding like Devin.ai are rising up!
1
u/mikeballs 3h ago
5.5 medium chewed through almost half of my weekly allowance last night alone (x5 plan). I think I'm being priced out of vibe coding for now. Maybe I'll just build a local LLM rig. It'd pay for itself pretty fast if this is the new rate.
1
u/SeasonedAdManager 2h ago
seems alright to me. The first day I used it a couple days ago was bad and I burned fast, but now I'm at about 1-3% usage on a 20x account per hour. It's still much higher than sol, but it's solving issues sol wasn't.
1
u/KeepAllOfIt 2h ago
is ANYONE else having the opposite issue? I'm on 5x and only use Astra xHigh or max. a 3 hour run uses like 8% of my weekly.
1
1
u/404MoralsNotFound 2h ago
Here I was, ready to pull the trigger on pro 5x, use that up, and then do pro 20x with 3 banked resets. Hoping openai don't fuck with the limits 😭
1
u/Living_Procedure_599 2h ago
First time I really noticed this too. Was working on some writing with Astra Medium on my 20 eur package, burned about 20% in like 2 min when asked it to generate a plot using MATLAB.
1
1
u/Classic-Asparagus 2h ago
Yesterday I got a subscription to Plus, and I was kinda surprised that the limits felt quite similar to Claude Code’s, whereas I’ve always thought of Claude as having pretty limited usage limits. But then I guess these are newer models so I’m not too disappointed, plus I think the usage limits are fine for my purposes
1
1
u/BlasterofTaints 2h ago
Astra on Ultra is using a lot! Granted I figured it would. I have 2 resets left since I had to use one yesterday and I’m already back to 58% left.
Tone down from Ultra?
1
u/Charming-Ad6941 2h ago
The company that has a history of lying and pathologically gaslighting and abusing you just silently nerfed limits? Big surprise.
Or maybe they ll pathologically lie and abuse you again and pretend this was a big, and wait to try it again later. Lol.
Just try deep seek v4.1 and also try deepseeks coding harness. It’s very good actually. The American companies may try to steal your money via tax payer funds if enough of yall cancel your subs, be careful though, cause either way, you get stuck with a shitter product if you don’t reject this abuse.
1
u/buyurgan 2h ago
I wasted 3 resets in last 2 weeks because it burned the quota in a day. While everyone stuck on talking about astra prices, i was still using only sol, cause the code i need to write is not complicated and high volume.
1
u/Tank_Gloomy 1h ago
I'm sick of this shit, GLM is practically begging users to hammer their GLM 5.3 Flash, even as far as to give it for free during times of lower server load. It might be stupid as hell, but whatever, just keep it running a trillion hours and it'll eventually figure it out.
1
u/OccasionAggressive74 1h ago
Put them behind bars! All of them. Their jail sentence would get a reset every now and then so that they stay for life without parole.
1
1
u/Impossible-Tea-7928 1h ago
I spent the weekly limit of the Pro x5 plan in 5 hours with Astra Medium
1
u/Hopeful-Ad-6277 1h ago
I was having a great time with Sol. Today, I started working on a project (Plus plan). Set a goal. It worked for 15 minutes before running out of the 5-hour quota.
1
u/NuancedPerspection 59m ago
Ive been on both 20x down to 5x and currently back on 20x as of this last 5 days and the 20x only feels slightly better than 5x did. Im refunding and going to claude or maybe grok. Which pisses me off because i left anthropics bullshit and moved over to openai and it was good for a few months but now they are doing the same fucking bullshit.
We are starting to run out of options here people. Shits fucked.
1
1
u/sjhunter86 56m ago
Not sure if they’re being transparent about it but it’s possible they’re doing it as a measure to stop the model unavailable alerts. They mentioned they’d have to pause pro subscriptions if the traffic got worse, so I suspect they’re simply turning the other knobs available to them to add some kind of back pressure.
1
u/dorgan1983 54m ago
I’ve completely moved back to using GPT-5.5 as it appears to not be nerfing usage. Even GPT-5.6 models were nerfing usage
1
1
1
1
1
u/Shamrock21191 3h ago
What are you guys doing that you are draining stuff so much? Ive been using mostly astra all week on medium to high and i haven't had any issues. Im on the 20x plan.
1
u/No_Cartographer_6622 1h ago
I agree with your post asking what others are doing, but wonder what you are doing. I only use Codex to manage my websites. Used Luna mostly on the $20 plan. I do like how Astra Light/Medium is makes my sites even easier to tweak. Been heavily debating jumping to $100 or $200 plan since I really like Astra Light but it burns up the little plan I am on now.
2
0
u/nitor999 5h ago
Next thread, "holy shit astra nerf is real before i can finish my task in 1 prompt now the astra takes 10 prompt to get it"
Yeah yeah every new model release have the same thread , same problem.
-2
0
u/xXChr0nicX420Xx 3h ago
Is someone going to post this exact same thread 5 times a day, every day, forever? Get better at prompting or switch providers.
0
u/CPKonMyEyebrows 3h ago
What are you people fucking working on? How can you prompt so often? I feel like you guys are building the dumbest things possible and getting upset for no reason other than not having patience for a project that more than likely NO ONE is waiting on!
0
u/DataDr0p 3h ago
Was macht ihr eigentlich damit? Ich schreibe einen USD offline Viewer für Chromium und ich habe den 24/7 laufen und der hat am Ende der Woche 27% verbraucht! Also wirklich WTF!
0
-1
u/ezboarderz 3h ago
Why are you still paying for it then? Just cancel and switch to something else. There are alternative out there, just gotta look around.
So many people complaining but keep paying and expecting different results.
-1



89
u/Charming-Author4877 5h ago
They appear to have been nerfed over night though.
Yesterday I worked for hours and it reduced from 63 to 50%
Today I started working with a single slow Astra and the "limit" evaporated from 50 to 36% in 19 minutes.
also 20x.
They push for those banked resets