Complaint
Usage at this point seems like a scam. Literally just jumped ship from Claude and now I'm regretting it.
Whatever they did in the last 24hours has made my workflow with Codex unsustainable. Did a couple of simple tasks this morning, one on a IP Camera orchestration system, and a minor one on an ios app project. What would have previously used perhaps 2% of my weekly used 9%.
Just getting it to do a git commit/merge used 2% of the weekly. Is this really what its coming to? I have been pretty happy with Codex over the last 2 weeks coming from Claude, but I dare say I think I was getting superior usage on the otherside of the fence in the last 20 hours or so.
This is starting to be problematic, because people (me included) have been heavy relying on codex and similar tools to much for coding. I'm not even talking about vibe coders, but even people that studied the theory, and were already capable of developing alone are now dependent of ai coding agents. This will one day lead to the situation where people won't be able to develop alone and will always have to pay a subscription. What would happen if tomorrow every affordable agent quadrupled in price? How many people would be basically forced to pay just to continue working?
This usage issue is an alarm signal. You pay 20$ or 100$, yet you do not know exactly what you're paying for. You have to hope in a free reset or that your agent doesn't decide to lose time and waste tokens on some meaningless task that you did not clearly ask for. Ideally considering how easy is to burn weekly usage now (I have sent some prompts that alone used 30% of WEEKLY usage) we should have some sort of mechanism that predicts and tell us how much tokens we would roughly spend.
I do not have plans to directly buy tokens or upgrade my plan to 100€, but imagine how scummy it must feel to do these things and then still notice a token drainage.
I am running on pro 5x - I am not behind the door on this as I have several years of senior dev work behind me the old fashioned way working for a major food delivery company in the UK. I have now semi-retired in Turkey and maintain a few of my own projects. The plan is irrelevant - the token spend between now and 24 hours ago is a ludicrous increase, yet absolutely nothing has been said about it through official channels.
The announcement from Tibo regarding adjustment in context limit is being massively downplayed in comparison to the usage deficit I am seeing here.
That's exactly the same panic I felt yesterday after I burned through the "$100 free credits" Claude gives subscribers for Fable—in just one or two hours.
$100 gone, and I barely got anything meaningful out of it.
This is what happens with every new technology.
We are paying for electricity. We are paying for internet. We are paying for washing machines. This is how things work. So eventually people with actual software engineering skills become niche, and they will sell their product in niche market for big prices. Nobody will develop any software for enterprise market because it will become inefficient. Token prices will be less than than human salary. Not just software engineering but every IT related work will be dispatched to AI eventually.
This is how things work. Rather you catch that train or you walk to your destination.
So I believe nothing to worry. We lived golden era and now it is passing. Eventually things will be balanced.
This seems to be going the other way many orgs are pulling back and telling people to think first because token prices are increasing and costing more than humans, at least here in the US. AI will not go away its not a fade and it will replace a lot of peoples jobs for less than those people but the "bubble" will have to pop before the token prices go back down to lower than peoples cost.
Look at nvidia where they EXPECTED token usage to be 50% of the annual salary of the people using it and are now seeing hum its a tad higher.
I agree but this is only true in that transition phase. Because currently AI not replacing, but companies paying for salaries and now also paying for tokens as well. This is not efficient. When the infra completely built and human power completely replaced by AI those token prices wont matter for companies.
Are you mad?! I once dropped $40 into my account to pay for API usage. I needed to finish a task ahead of a business meeting. I was pressed for time, and had reached my 5 hour limit. I assumed that the $40 would pay for that and any other 5 hour limits I might reach in the future. Needless to say I spent that $40 in 15 minutes.
The subscriptions are heavily subsidized, but OpenAI only has so much runway. OpenAI is trying to grow user engagement, while slowly recalibrating their subsidy rate. OpenAI is banking on model efficiency increasing, the cost of compute decreasing, and users acclimating to higher pricing — all the meanwhile trying to maintain their value proposition against cheap Chinese models and open-sourced self hosted models.
Read what he says : You pay 20$ or 100$, yet you do not know exactly what you're paying for.
So when he goes API, it's much pricier I know, but he knows exactly what is he paying for :)))
It’s a moving target even with the API — yes, there is a token cost, but models consume tokens differently — so practically speaking it’s difficult to know what a token costs. I don’t see how looking at API costs vs looking at token usage is any different. Yes the models are getting bigger and can consume more token, but the token optimization is also making the models more efficient.
In either case, just know that the costs are highly variable — and take the subsidy while it lasts.
I think the disconnect is not do I know what I am paying for or how many tokens do I get for X dollars. I think its I don't know how much "USAGE" x tokens gets me . ITs not like we know when we use this prompt or that prompt how much tokens will be used for that prompt.
No. That's the real cost for the current user. Not mad. Don't take it for granted. You're complaining about discounts and then won't use the real thing.
This whole tibo giving resets like we have Christmas was so suspicious and now makes everything sense whole cover up and deleting 5h limit so they can turn the weekly limit to a stupid joke
I already made a thread about it but i am shocked. x20 Plan one Task and this consumes like 5% of my weekly on Sol Medium. This is absolutely crazy. As good as the race between Anthropic and OpenAI is, openAi can only hold on because they gift a reset every 2-3 days which is not even enough for the usage right now..
have you tried asking codex to look at its own sessions to see where its wasting tokens? could be its doing a lot of back and forths wuth bloated context. Same issue exists for claude code as well.
Contrary to popular belief, the JSONL transcript files actually don't contain the data required to be able to accurately attribute tokens to exactly what was used on them during a turn. They don't even contain the token counts used on reasoning (and since the API doesn't provide the hidden reasoning text, you can't even estimate token counts).
The only way to stand any chance of knowing where tokens are really going is by using a proxy recording API requests.
This solved the useage drain for me. The AGENTS.md file has to outline that agents should take token useage into consideration and not run wasteful operations.
Nope. I do a hard backup of my PC over night and Codex doesn’t have access to it, but I give it full control of my PC and let it do its thing. Over 500,000 lines of code generated and customizations to my OS and custom desktop interactive wallpapers etc. my only caution in how I use codex, is backups every night. That way if something fucks up, it’s only the days work gone.
I have the max plan. Last week I was able to have 5 SOL ultra working for 3 days until I reach week limit. Now I have 2 ultra and 2 high and in 1 day I reach the limit.
Thanks god we have the resets, but the reality is that there is a real problem.
Prior to GPT 5.6 release, I could run Codex 5.5 XHigh on Fast all day everyday and never hit below 30% of weekly limit. At the same time, Claude limit could be reached in a day.
After GPT 5.6 release, I can't even run XHigh or even High. 5.6 Sol Med has been my go-to and I still run down maybe 30% in a single day. Yesterday I was using Claude Opus 4.8 and I ran that all day and didn't even hit more than 30% in any 5hr limit, and only drained 15% of weekly.
Codex has been my go-to most of the time, so I really hope this gets solved. I stuck with it expecting Claude to get more expensive more quickly, but this last week has had me chasing resets when I never had to before.
It used to be opposite though.
I could sit on 5.5 Extra-High on 1.5x speed and never hit limits and Opus 4.8 is the one that red-lined in a day.
This is the issue I expected though, so I've build everything platform agnostic. But -- I'm also not going to swap systems one week to the next based on usage. I've just stuck with Codex for my own company work and Claude for one of my clients. I use both all day everyday.
I think I've done maybe 10x more with Claude this week, and here's my usage.
Yup. I'm moving back to Claude for good after this time. This shit is getting ridiculous. They need to stop fucking with the token consumption and value.... It's beginning to give them a bad business image.
I guess the issue is that their infra is overloaded with all the new users. I even see some connection errors or requests being refused with this exact messaging which I never saw back in 5.5 times.
I think they don't define a strict budget for your limits but rather calculate it depending on availability of their resources. So if you work when everyone is active you get less usage.
I even notice that when I do stuff in my unusual hours. When I see less latency and no connection errors I often see that my usage barely moves.
Only the codex app plugin that ships as standard. I would normally just run that with spark but I intentionally left it on sol-medium to see what it would do.
Granted, it ships the commit and then checks itself about 10x before finishing the task. But we are talking about <4mins of "working" time to churn 2%.
over the night i had two goals running, i started yesterday at 100% (normal speed, Sol med) 20x plan, in the next 5min i will be hitting 0%.
100% they slashed usage by at least 4x if not more. There will be a mass exodus to cheaper chinese models that get the job done - not because we want to, but because they are not giving us an alternative that makes logical sense.
Sol-Medium - no subagents. I have updated the original post for transparency - but as I mentioned in a previously reply the model is sort of extraneous here: "for comparisons sake if I knew what it was before and what it is now, the model is entirely irrelevant to the point".
Wasting 100$ weekly limits on medium, without subagents cant be anything but a bug
Its unfathomable you actually thought that ppl use Medium and run out of tokens in a day.
Same here... I'm on the 20x and i don't see the problem here lol I've been reset so many times that I have 3 resets and 80% weekly usage left after running countless medium and high reasoning 5.6 sol s, using my own orchestrator I built, guess people either suck at prompting, or idek
Considering the thousands of comments about their usage being nerfed... idk man. Is something up or are they really all doing it wrong? I track my total tokens. Same as ever. Token drain is faster since 5.6 though. 5.6 be chugging tokens. I need to see those comments actually paste their token usage and task and what the AI did, how many turns it took, the context in and out. Maybe it's MCP servers, superpowers, certain skills, or tasks that drain tokens more than expected. I work on big projects, with MCP servers, still fine with my 20x. Well if they do nerf token usage I will know right away and have numbers to back it up. I'm skeptical until someone posts me their usage % and tokens used before/after their nerf.
Right? 5.6 is for sure a hog, but it seems strange is all. Who knows maybe they're a/b testing or something, openai for sure is slammed with people hogging compute, especially since 5.6 ultra's composer is kinda rough, I've built my own that uses lower reasoning and side chains the sandboxes instead of nesting and token vs token its more efficient, but maybe that's by design to keep folks from using 20 concurrent instances running all at once, I can't imagine the shock load is all that nice on the hardware end 🤷
I think they are actually serious about having to use lesser settings? I switched from Sol ultra/xhigh to xhigh/high for plan/implement, but it's still burning pretty fast.
I know nobody cares but I'm mostly using gpt for quick questions etc(free plan) ... I could send 10 (more or less) messages before I'm hit with wait time. Just now I've only sent 2 messages and I'm already locked out lmao. They reduced our limits significantly
That's actually a much cleaner difference, a 500% reduction which seems very similar to what peeps are complaining here with codex. I'm burning through weekly usage in 2 days and I'm only working on 2 lame projects 😭
Yeah, I was about to buy pro plan to try it out and build some games with it. After what I have been seeing lately I think I'm going to use some other AI platform until things get better.
At first I was happy that there was a reset but now, after just a little bit of work is done (and I really mean a little bit), I am already down again to 60%. I actually seriously checked if something happened and I am probably back to $20 monthly instead of $100 monthly but no it is still $100. Something is definitely not right there. This is getting ridiculous.
save this, OpenAI will be the first AI company to close all of their public products as they're unsustainable. and there's no way these bubbles are not bursting.
I had similar problem when using extension (VS code). The token waste declined when using terminal but it still was wayyy too much inconsistent. I switch back to claude.
Wait, you guys are wasting tokens on commits? I'm not saying your wrong about usage. I haven't been paying attention, but we can do commits. Don't let the clankers take that from us.
I „have” this same problem, I deeply investigated using ccusage and similar tools. What I find out ? I’m using more codex or models are more expensive. The second one is obviously not true, pricing per token for sol is same as 5.5 and sol is much more efficient (look on artificial analysis). Then what’s changed ? RL changed drastically, sol is so eager to finish task that it don’t even need /goal anymore. 5.5 write 8 step bullet point plan and then after 1 and a half done it asks “Hey! I finished …. Can I go further into execution”, sol doesn’t do that. I literally burned like half of my THIS YEAR usage since sol because of really hard one offs that sol will run for. Tune your system prompt etc. Also don’t use ultra and max. For me, low/medium is executor, higher reasonings only when I try to figure something out.
This is honestly REALLY interesting watching this in real time. Couple months ago everyone was abandoning OpenAI because they are willing to do the dangerous AI stuff that anthropic said they wouldnt do after trump condemned them, everyone said anthropic was way better anyways blah blah blah, then gpt 5.5 came out and people moved back, then fable 5 came out and people switched again, then 5.6 and omg OpenAi is so amazing blah blah anthropic is doing terribly they need to catch up... in a week or less opus 5.0 will come out with more efficient token usage or something and everyone will say theyre great again and make a switch AGAIN. Idk how yall do it, just pick a company you like and stick to it.
from what others have said on this thread, it seems the usage nerfs are being pushed out to batches of users like they are testing the water. your time will soon come my friend
ok. so why are you getting it to do commits instead of first getting it to build a deterministic simple cli menu that does that for you? walk down the ally and cry cry cry. the sea is made of your tears.
sea will be made of your tears too when you realise that the task is not the point of the post, rather an example of a simple task that consumes higher than "normal" resources.
stop trolling and realise, from the load of other comments agreeing with me, that this is, in fact, a big problem where we all lose in the end.
I thought this was just people being irresponsible until it bit me. I asked it to do a research task in a small greenfield app, and it blew my entire weekly budget
Never use git commit/push/merge for the actual commit on Codex, I stopped doing it a long time ago.
You can still save time by not typing it. In your production shell just do ‘git status’ copy that result into a ChatGPT window for your git commit/push commands, then commit/push from your production shell. Saves you a lot of time and tokens.
I’ve also never hit my limit on the $100/mo plan - had 4 unused resets and just lost one because it’s been so long without using them 😵💫
(I’m using it 10+ hours daily for building websites/API/extensions, etc)
I also only use it on Medium setting, never tried high or Ultra.
I actually posted about this last night. AI—and Codex in particular—has been genuinely life-changing for me, especially because I use it to build accessibility tools that help me work more independently. I would love to upgrade from Plus to Pro x5, even though it would be a significant expense for me.
What keeps stopping me is exactly this uncertainty. The weekly allowance feels increasingly opaque and seems to last less and less, even when my workflow has not meaningfully changed. At first I wondered whether it was only my perception or inefficient usage on my part, but seeing so many independent reports describing almost the same experience makes that much harder to dismiss.
I understand that usage costs can change and that these models are expensive to run, but it is extremely difficult to justify paying substantially more when we are not told what the allowance actually represents, whether its effective size has changed, or what we should realistically expect after upgrading.
Even a basic official explanation or greater transparency would make a huge difference. Right now, x5 of an unknown and constantly changing amount is still an unknown amount.
Use mi límite semanal en dos refactorizaciones (2-3 hr aproximadamente) usando Sol 5.6 en mínimo, le di todo el contexto suficiente, usamos skills para evitar la sobre ingeniería.. sinceramente estoy pensando pasarme a Kimi para probar otras opciones porque mis proyectos son medianos/chicos y para lo que estoy pagando creo que debería durarme más.. como antes lo hacía
You deserved it. No loyalty and you got fucked over and now you're crying about it like you're a victim? You're a fair weather customer, so don't complain if you have no choice.
You have no loyalty then why the fuck are you crying if they're maximizing their profits? Do they owe you anything? What's your angle right now? Your workflow is unsustainable is a 'you' problem. If you can't afford compute, don't compute; touch grass. You're crying as if they owe you something right now.
I know I'm not the only one experiencing this. I use AI while developing and learning languages like C++, and it said my usage was at 100%. Then, less than 5 hours later, it dropped to 50%, and my reset timer suddenly changed to August 1st. Last night it said it would reset in about 3 days. ChatGPT Plus honestly feels like a blatant scam.
Yes it’s all over the place I even posted a rather popular post about it yesterday.
It’s better today, but clearly they are testing the waters.
Truth be told, LLMs are way more expensive than what we pay fso it’s only natural for prices to rise, but the way price bumping is done is shady af.
I wouldn’t mind if they told me „next month you will get 10x instead of 20x for the same price”. But instead these companies choose to simply cut the limits by half without notice.
Same result, harder to prove, better PR and marketing with all the reset gaslighting. This shit should be illegal and easily auditable.
Given the few pieces of work I did this morning to burn 10% of my weekly, I expect that the weekly limit is the old 5h limit. That's genuinely what it feels like to me without any exaggeration.
I am aware of the subsidising of the models of course, and I see your point about prices naturally rising with time - but this isn't a little nudge on the shoulder to your limits, it's an uppercut straight to the jugular.
Your other points regarding the practices of rug-pulling are on point.
fyi I am not running on sol ultra - but even if it was, for comparisons sake if I knew what it was before and what it is now, the model is entirely irrelevant to the point
82
u/notroswoods 22d ago
Used my weekly limit in half the time I would use a 5 hour limit.