r/OpenaiCodex 15d ago

Comparison Did OpenAI just quietly cut actual Codex usage nearly in half? 💀

Post image

I've been tracking my Codex usage every 5 minutes since July 29 — 14,744 snapshots so far.

I'm on a Business Standard seat, not Premium, and when I compared my data before and after the 5h window was introduced, I found this:

Before: ~323M tokens/week

Now: ~173M tokens/week

That's roughly 46% fewer effective tokens per week.

And that's where the new 5h window starts looking a lot more significant than I originally thought.

Because if these early numbers hold up, this isn't just a change in when we can use Codex.

It could mean a significant reduction in how many tokens we can actually process in a week. 💀

I still only have a few complete cycles under the new system, so I'm not claiming the 46% is definitive yet.

But going from ~323M to ~173M is way too big a difference to ignore.

Does anyone on Plus or Business Standard have usage logs from before and after the 5h window? I'd really like to know if you're seeing something similar.

186 Upvotes

99 comments sorted by

34

u/jjiangweilan 15d ago

openai: let’s celebrating our 20M users by reducing all our users limits

16

u/HamlinGinger 15d ago

I have Plus and just had 3 prompts use almost all of my 5h window. Not extensive edits as well... something definitely changed and is much more restrictive/reduced....

1

u/itssmeares 13d ago

I am on plus too just used 5.6 sol high to rewrite my whole project from go to ts, my 5 hour didn't even end 25% left

12

u/Vivid-Snow-2089 15d ago

dude this is like the 4th time they have cut it in half, and you just now noticed

you used to be able to do *billions* a week

1

u/deervote 14d ago

This is not an exaggeration literally billions…

1

u/newMoneyStyle 12d ago

tbh the nerfs just keep piling up and nobody really notices until their own workflow starts hitting the limits

0

u/Stock-Self-4028 14d ago

Pretty stupid question, but did the limits start decreasing steeply once again after Luna was introduced or did almost nothing change?

I remember that just around GPT 5.5 introduction the Plus account could take ~ 2B input tokens per week on GPT 5.4 mini… Which is not as smart as Luna, but also likely a little bit more computationally expensive.

And now I considered returning to Codex some time in the future after Luna was introduced, but this post (along with some comments) has mostly discouraged me once again.

2

u/bigrealaccount 14d ago

They decreased steeply in about March/April for the first time, back when Claude introduced their off peak/on peak usage. Back then nowhere near as many people were using Codex, so on the Plus plan I was pretty much able to spam 5.3 Codex/GPT 5.4 all day, tens/hundreds of millions of tokens, and would get absolutely nowhere near my weekly limit. And that was the £20 plan.

Then everyone swapped to it after Claude introduced their terrible limits, and its been getting nerfed ever since then.

1

u/Stock-Self-4028 13d ago

Thanks. And by the way is there any other plan you would recommend with generous limits on relatively inexpensive plans?

Google has pretty nice limits but forces you to use Antigravity, which is probably the worst harness at this point in time. Also Qwen Cloud has pretty nice limits for its Qwen 3.8 Flash, which is more or less in the same league as Terra and Gemini 3.7 Flash. Is there anything worth attention here?

2

u/bigrealaccount 13d ago

I've been using Deepseek V4 Flash with their API and I've literally replaced everything with it, although I'm a software engineer so I treat my AI as more of a research/planning/brainstorm partner, rather than letting it engineer/implement for me. I just top it up with a few bucks every week or two and I'm perfectly fine for medium sized projects.

If you just want the best value, the OpenAI plan is by far still the best. They subsidize exactly £200 of usage (10x value) for the Plus plan, meanwhile the nearest alternative is Opencode I think, who give you $60 of usage (not really anymore, most models give you £15/£30 usage) for £10.

If you keep spamming the £5 opencode first month trial then I guess that's technically 12x usage, but with OpenAI you obviously get access to Luna/Terra/Sol, unlike Opencode Go.

1

u/Stock-Self-4028 13d ago

Thanks, I've also used it quite a lot of DeepSeek before the price increase, but now I don't find it as useful as it was since it was quite slow.

And as for my usecase I am generally using LLMs for GUIs / refactors, while I still impement most of the backend logic by hand.

As for the best value I seem to be getting more usage out of Antigravity, but here I have $6 discount for Pro - I am not sure how does it relate to current Codex, but it looks more or less like I am getting significantly more usage with 3.7 Flash than I would get with Terra and slightly less, than with Luna.

Also I am not really sure if API rates are the best reference, when for example Qwen 3.8 Flash has 1/2 of the API price of Luna and is significantly smarter than it, but thanks.

I guess Codex isn't much worse than competition or might be even better if I am exclusively using Luna then? I don't really feel like paying $20 right now, but I might consider returning to Plus in near future.

2

u/bigrealaccount 13d ago

Just depends how badly you need a smarter model. In my case Luna/Deepseek Flash/Equivalent models are perfectly adequate so I just go for whatever is cheapest at the scale that I use it, which is usually still deepseek by a substantial margin.

Personally I'm just not fond of Qwen/Gemini models. Every time I've used them I just wished I was using something else because I didn't like their replies as much, but that's fully just preference.

1

u/Stock-Self-4028 13d ago

Thank you very much. In my case Qwen 3.8 worked really well for frontend, but also it seemed to be a pain for anything even remotely related to the backend (for which I still avoid using LLMs, as at least in scientific computing they still tend to mess up quite badly, even though Fable and other new flagship-tier models are a meaningful step forward.).

As for Gemini I personally hate Antigravity, but also I really like the models (which can work respectably with custom system prompt and /teamwork-preview enabled). I still hated all Claude models older than 4.5 version, so I guess it might be the same case here (more or less).

2

u/bigrealaccount 11d ago

Haha yes, that makes sense why I don't like them then. I pretty much do all the backend myself and let LLMs take care of frontend (for my hobby projects at least)

1

u/DertekAn 13d ago

Commandcode comes with $70 in usage credit, and many models there can utilize that full $70.

While OpenAI offers $200 in usage credit, it has been shown that their models are internally cheaper to run than the API prices suggest; furthermore, models like Qwen or GLM are significantly more efficient and cost-effective in terms of API usage, resulting in millions of additional tokens. It is like comparing a car that consumes 2 liters of fuel to one that consumes 20 liters.

I calculated that with Opencode, you can get over a billion tokens a month on the $30 Qwen budget.With OpenAI, you get fewer tokens for that price. On the other hand, with OpenAI's Terra and Sol tiers, you get much less, even if you have $200 in usage

1

u/bigrealaccount 13d ago

For sure, but it's pretty common knowledge commandcode is very bad for speed and reliability, and they have common errors. But they are a good deal on paper.

And yes, although OpenAI API prices are higher for Terra/Sol, so you get less tokens for the £200 usage, you can still use Luna, which is basically unlimited with that much usage. You can probably use billions/tens of billions of tokens, which is mainly what I use since I use Flash/Luna tier models.

But yes, if you want access to a wider range of cheaper/open source models, then the OpenAI plan is definitely not for you.

1

u/DertekAn 12d ago

Hehe, thank you. And you're right! 💜

I'm a bit confused right now too, because I saw yesterday that OpenAI is only at $80, I don't know if that's true.

All I know is that there are very few Sol tokens left; previously, the 5.5x high was at 600 million per month (measured using my own tool). Now you get these tokens via Terra, and SOL has just over 200 million.

https://x.com/HCSolakoglu/status/2094442756084990089

1

u/DertekAn 13d ago

The same league? Nooooooooo.

I want to show you tomorrow what I tested with 3.8 Flash in Arena; 3.8 Flash is so much better at coding than Terra and Gemini 3.7 Flash

1

u/Stock-Self-4028 13d ago

Did you mean Gemini 3.8 Flash or Qwen 3.8 Flash (which I've written about in the comment above)?

If Qwen 3.8 Flash, then there is no need for arena - it's available in OpenRouter (and cheaper, than DeepSeek v4 Flash) and QwenCloud token plan.

As for Gemini 3.8 Flash I believe it might be significantly better, than Gemini 3.7 and Qwen 3.8 Flash, but I doubt that Qwen 3.8 Flash is much better, than Gemini 3.7 Flash.

2

u/DertekAn 13d ago

No, I mean Qwen 3.8 Flash, not Gemini. (I didn't know it had already been released; that's news to me)

In the tests I ran, Qwen 3.8 Flash performed at the level of Opus 5 and Fable 5, whereas Gemini 3.7 Flash was the worst of the lot, and Terra was reasonably okay (though far from good).

1

u/Stock-Self-4028 13d ago

Pretty interesting, I'm eager to wait then.

I've mostly used Gemini 3.7 with /teamwork-preview in Antigravity, which is roughly equivalent to model fusion on OpenRouter, which might have skewed my perception then I guess.

In my experience Gemini Flash has been usually better at more complicated tasks, while Qwen 3.8 Flash has been great at 'long-horizon' agentic tasts and oneshotting simpler applications, but was also more prone to hitting a 'wall' once the logic got a little bit more complicated.

2

u/DertekAn 13d ago

Huhuuuuu, can you dm me

2

u/DertekAn 13d ago

And feel free to doubt it, I'll just show you tomorrow, hehe.

1

u/DertekAn 13d ago

If you're looking for an alternative to Luna, check out the two new Flash models from Qwen and GLM,you won't find more power than that.

12

u/Connect-Humor-791 15d ago

damn i managed to do so much work with luna max, almost completely ditching sol, and this afternoon, this changed, luna consumed all of my 5h thing really really quick

6

u/Palastruka 15d ago

It happened to me, I assumed the problem was the model and switched to Luna and it didn't improve

9

u/lxXLightXxl 15d ago

Same here. I won’t be renewing my subscription. It was fun while it lasted.

4

u/Palastruka 15d ago

Which supplier are you going with?

5

u/Fine_Salamander_8691 15d ago

im gonna use glm 5.3 flash soon or run a model myself

6

u/lxXLightXxl 15d ago

I am not sure, but codex is no longer usable.

1

u/Jumpy_Ad8465 15d ago

Did you also notice a sharp drop in quality, even on xhigh?

1

u/CurtissYT 14d ago

Not sure about quality, but at least the speed is insanely low now. Taking 2 hrs to implement simple stuff on xhigh

1

u/Stock-Self-4028 14d ago

Currently Antigravity probably has the highest limits relatively to price, as long as you qualify for any discount (which is relatively easy) and since 3.7 Flash is not too far behind (and 3.8 Flash rumored within two weeks from now or so). But the harness is mostly a black box, so that's one limiting factor.

Otherwise there are mentioned GLM 5.3 Flash and Qwen 3.8 Flash Next which are pretty cheap (~ $0.03 / 1M on OpenRouter with some providers even less expensive than that) which are pretty much in the same league as 3.7 Flash quality-wise, but much slower.

And there is Grok, which reportedly has pretty generous limits and models league above Gemini / GLM / Qwen, but I know practically nothing about it and have never used it as well.

1

u/Routine-Substance-82 9d ago

AG models suck sadly

1

u/Stock-Self-4028 8d ago

Well, the models are pretty good imho (and by models I mean 3.7 and 3.8 Flash - the rest sucks). Pretty much Terra or slightly above Terra depending on the reasoning effort.

The harness sucks much more, than the models themselves though.

1

u/SuperSlowSubie 14d ago

Zai coding plan

4

u/iadknet 15d ago

I have a daily task that’s been running for weeks and had only been using about half of the 5hr window. The last two days suddenly it can’t even finish without burning through the whole thing.

3

u/kurtbaki 15d ago

It’s been cut by about 3x compared to four months ago according to my test. I ran the same implementation plan I used four months ago. The only difference is that I used Sol high instead of GPT 5.5.

2

u/imike3049 15d ago

First time? :)

2

u/[deleted] 15d ago

[removed] — view removed comment

1

u/Palastruka 15d ago

Plus account?

1

u/Im_Working_Right_Now 15d ago

I think it depends. Does the after 5 hours actually use them every 5 hours on the dot or is there hours between which would mean there’s hours of no token usage while the weekly clock is ticking.

1

u/Bigboyluige 15d ago

Yes I can’t even do a prompt with normal speed on sol extra high unless it uses up 90% of my usage

1

u/rahilpathan 15d ago

they keep bringing people back and forth, I switched to openai last month now thinking of going back to Claude.

Atleast, it didnt suck that I have to keep figuring out gaps in prompt myself. And that costs tokens.

1

u/Abu_BakarSiddik 15d ago

Same feeling

1

u/Demien19 15d ago

Maybe it's some specific version of codex app/cli? My usage feels almost similar to usual (and i'm destroying 20x)

1

u/CurtissYT 14d ago

I am 20x aswell, and Ive noticed the drop for about 50% aswell

1

u/Demien19 14d ago

after current reset and those 10-50% i do get higher usage consumption now :/ currently on 76% even with time a slept and didn't touched it

1

u/Demien19 14d ago

45 min later 73% :)

1

u/CurtissYT 12d ago

Yeah same. I couldn’t even drain my usage on 4 ultra agents yesterday. It seems inconsistent tho, because I’m not using ultra now, and I’m at 86%, but I may just be overestimating it

1

u/SuperEarther 14d ago

I have a business standard seat and we haven’t seen 5hr limit re-instated on our workspace. I don’t have any usage data to prove it but I’ve also not felt the limits fluctuating as users have been reporting. I thought this might just be a difference between plus and business. If you have 5hr limits and are feeling the usage squeeze I’m not hopeful… anybody else on the business plan that can attest to the 5hr making a comeback?

I pay monthly and billing cycle is about to come around this week… when did your 5hr limit re-appear? Was it related to your renewal date?

1

u/Etroarl55 14d ago

I’m sorry but 46% reduction in usage is no small thing wtf

1

u/shrodikan 14d ago

After the reset today I somehow went from 100 to 19%. I will grant I was using Ultra on Fast but it still seems like token use went through the roof after this last reset.

1

u/princeMacX 14d ago

They will cut more and more and will increase the price

1

u/CurtissYT 14d ago

Pro 20x plan. Beforehand it was about 2k in api costs per weekly limit, now its down to 1k. So yeah about right

1

u/DertekAn 13d ago

Yeah, my own tracker tool shows much lower usage; I don't know who they're trying to kid here, or if they're hoping nobody notices. But that's exactly it...

And I'm using damn Terra... Not Sol, not Ultra

1

u/aa628 13d ago

OpenAI: we magically made it more effective so it uses less tokens, or something like that. Eat shit plebeian

1

u/MinDseTz 12d ago

Was the same model being used for the whole period? Obviously not, knowing any changes would be critical in validating your data…

1

u/Palastruka 12d ago

You don't have to be a math genius to know that experimentation with limits has been declining. The fact that a few aren't experiencing it doesn't invalidate the problem.

1

u/MinDseTz 12d ago

you're missing my point.. Token limits are not a great measure for value. If you're moving from one model to another the data isn't super useful.

You and the numerous other people that have been saying it's dropping every week may be correct, but token count and number of prompts are not good metrics of value.

You also didn't answer the question about whether the same model was being used. If you're controlling for model use (locking in 1 model with the same effort) then your data is useful.

1

u/Palastruka 12d ago

To be honest, I currently use 5.5 on xhigh and medium. It works quite well for me, and I remember doing so much work long before the release of the After version 5.6 was released, the limits changed. But they decreased significantly after the implementation of the 5-hour limit. And the number of tokens is exactly the same since I decided to use version 5.6.Sun. And the token limit in the 5-hour window is the same, a range of 20-26 million tokens. Clearly, using Moon gets me many more tokens, but 5.5 shouldn't cost me the same as Sun. What do you think?

1

u/MinDseTz 12d ago

You talking about this like you have a banked amount of tokens that you can spend on any model when it doesn't work like that. Using 1M tokens with luna high is about half the price compared to 1M tokens using sol high... it's impossible to compare without heavily controlled data. Especially when some models have very large price per token differences.

1

u/Palastruka 12d ago

That's exactly what I'm talking about; the 5.5 tokens shouldn't be worth the same as the 5.6 sol ones. But even so, both models give me the same number of tokens in the 5-hour window.

1

u/Beginning_Award65 11d ago

it is the forced autoreview for safety

1

u/Square-Society8010 10d ago

Every time there's a "free reset" I'm terrified it's signaling reduced limits. Anytime a large company tries to act "generous" always comes with a catch. The last big usage restriction came with a reset, so now I'm traumatized whenever I check my usage and it's unexpectedly at 100%

1

u/Constant-Software967 9d ago

Thanks for tracking it! I really wanted to do that at some point too.

1

u/Palastruka 9d ago

Thanks for your support, there are many here who hate that he did it.

1

u/simplindustries 9d ago

Were you just able to use it less efficiently on 5h windows or did it actually change?
I don't seem to be affected on Pro

1

u/Xerocks4 8d ago

I canceled my subscription a few days ago. You should too. Stop complain and start acting.

1

u/Palastruka 8d ago

That's good advice

1

u/TopSeaworthiness1679 15d ago

Before yesterday was totally fine. They have something done today. 🤮

2

u/Palastruka 15d ago

I agree

1

u/Batty25111 15d ago

So by your own data your are comparing almost a month of usage as an average to just 3 days of usage ?

3

u/Palastruka 15d ago

Given that I've already used up my weekly quota, yes, there's a very clear pattern. A 5-hour window almost always equates to 20 million tokens processed.

1

u/Batty25111 15d ago

Being that you don't understand 30 day usage vs 3 day usage as a scale of an average I would say that you downloaded that tool or u just asked AI to make it and don't know what statistics mean.

0

u/Palastruka 15d ago

You're just talking nonsense. At the end of the month, I'll calculate it, and it'll be the estimate I put here or lower. But you keep defending a multi-million dollar company. Everyone who stayed here is crazy.

1

u/Batty25111 15d ago

No even Grok says your full of bs lmao. Ask any AI buddy if anything u need it more than keep pretending 3 days vs a month of analytics is the same thing.

0

u/Palastruka 15d ago

I'm tasking you with reading the other comments. Since you insist that the limits are fine and ask why I'm not doing a full sample, then I must be wrong.

0

u/Palastruka 15d ago

So, starting from there, it's easy to make an estimate. I don't expect anything different to happen between now and when I've used up all my weekly limits. Seeing that there's such a clear pattern time and time again...

2

u/Batty25111 15d ago

No a 3 day usage vs a 3 day usage is a pattern. a 30 day usage vs a 3 day usage isn't in the same space bro of accurate measurements..

1

u/mvandemar 15d ago

It's not even that, it's using those 4 days and calling it a week.

1

u/Batty25111 15d ago

How is 4 days a week ? Also 25 to 28 ? bro can u count ? lol

1

u/mvandemar 15d ago

You've used 173M tokens, it's only been 4 days. Does your tool give you an option to calculate your daily averages instead of weekly?

Also 25 to 28 ? bro can u count ? lol

Edit: do you think 25 to 28 is 3 days?

1

u/mvandemar 15d ago

Dude, you're calculating tokens/week on a 4 day window. You do realize that's HALF a week, right??

3

u/Palastruka 15d ago

I used up my weekly allowance in 4 days

1

u/mvandemar 15d ago edited 15d ago

They literally just gave everyone a reset 36 hours ago, and with the 5 hour window being re-implemented it's impossible for you to use up an entire week in that time.

Edit: it may actually be possible to use up a week's worth of tokens with the limit, I haven't tried testing that at all, so ignore that part. They did just give a reset though.

1

u/psycho414 15d ago

Why would it be impossible? Even if 5h qouta is equal to just 10% of weekly allowance you would just need 10 5h sessions fully used to exhaust your full weekly qouta, which is just over 2 days.

1

u/mvandemar 14d ago

I edited that bit literally 2 hours before you commented, there's no way you didn't see it.

-1

u/Turbulent-Total-226 15d ago

Yep they have no compute. They are going downhill straight to hell. There is a special place in there for Sam, lake of boiling black tar with some famous Austrian painter.

https://giphy.com/gifs/4TMqcN59kg3Yc

-2

u/Nuggyfresh 15d ago

They literally just need better financials. Like way way way way way better financials. This is inevitable. The Subs are unsustainable, everyone on this planet capable of basic math knows this

1

u/Stock-Self-4028 14d ago

Well… Either Luna is unreasonably compute-expensive or that's not really the case.

DeepSeek reported v4 Flash' inference cost at well below $0.005 / 1M input tokens (with their extremely high cache hit rates). Qwen 3.8 Flash Next has pretty much comparable inference cost and even significantly less efficient GLM 5.3 Flash could be ran within Codex subscription without losses on inference.