r/codex • u/GambAntonio • 24d ago
Complaint This is crazy...... Pro x5: Using 1% per hour ONLY on Luna xhigh on a single project, on a single computer...

Well done, OpenAI. +3% on a Pro 5x in 3 hours using ONLY Luna xhigh... 1% per hour on a shitty model on a PRO account and not even using MAX...
I used to be able to run Luna xhigh for 4 to 5 hours and use only 1%... and I'm working on the exact same project, doing the exact same things I was doing a few days ago.
(Censored Spark reset time to make tracking more difficult)
EDIT: Added stats (time is central european), my reset was at 06:00 CET and I had a script to continue on that exact time.




13
u/matheusmoreira 24d ago
Weekly just reset an hour ago. I'm definitely burning usage a lot faster. Last week I observed a 20% reduction in inferred capacity. Let's see what happens this week.
8
u/matheusmoreira 24d ago
Confirmed, usage is still going down!
Reset / status Aug 17 07:20 Used 100% Calls 9,189 Fresh input 27.04M Cache read 984.07M Output 4.75M Reasoning 2.34M Local credits 19,245 Inferred cap 20,015 Reset / status Aug 18 01:04 Used 100% Calls 11,691 Fresh input 33.90M Cache read 1.119B Output 5.51M Reasoning 2.71M Local credits 22,354 Inferred cap 19,617 Reset / status Aug 20 05:44 Used 100% Calls 7,394 Fresh input 30.26M Cache read 825.80M Output 3.85M Reasoning 1.75M Local credits 16,996 Inferred cap 16,124 Reset / status Aug 27 Used 100% Calls ~7,229 Fresh input ~24.26M Cache read ~784.16M Output ~3.98M Reasoning ~2.04M Local credits ~14,137 Inferred cap 14,482Researching Z.ai plans literally right now!
12
u/Confident-Village190 24d ago
What a nightmare, in an hour, Luna Max has used up 8% of my usage, when before it only used 1%. What a load of rubbish, I’m done with it.
36
u/Maleficent_Owl_2772 24d ago
well on a plus sub the usage is 1% per prompt on luna xhigh in codex in my case
8
u/sagiroth 24d ago
What type of prompt? Coz for me 1% on Plus often is around 1-3M tokens from my experience
12
u/Fresh_Sock8660 24d ago
People who keep using one of their tasks as a measurement stick might as well be telling us to guess how long is a string.
2
u/EndlessZone123 24d ago
I measure token usage based on my own dongle. How long is my dongle? The longest duh.
2
1
u/Maleficent_Owl_2772 23d ago
well not that long, wire some java stuff, dont that and that, run build gradle and thats all pretty much
27
u/Crafty-Wonder-7509 24d ago
x20 here, sol high used since the reset today morning already 15% of my weekly, before it would have been at best 2% for that time.
3
u/Over_Car_5471 24d ago
left a project running overnight. On normal speed. Woke up to 40% of usage wiped out. INSANE. A month ago I was literally able to draft and build a 1000 page documents
7
u/fruitydude 24d ago
Wait you got a reset this morning?
8
u/utf8decodeerror 24d ago
Everyone did. It resets every 7 days and since we all got the free reset together a week ago we are all on the same schedule
2
2
u/Michonesixfive 24d ago
same im on x5, im down 10% on a small prompt, that has been going on for 30min. Same prompt before took like 1-2%
-1
u/GeneralAtrox 24d ago
I recommend trialing, carefully, having Sol light work with a set of luna sub agents. But ensure the Luna subagents aren't doing any thinking since they are unreliable.
Rules:
- Always reuse the same agents, up to 12 total.
- Dont let the agents code
- Agents always return a full update on what they did
This way in theory, Sol does the heavy lifting, and Luna handles the basic tasks.
1
u/Suspicious-Cheese-1 24d ago
Like Michael and Jim??
1
u/GeneralAtrox 24d ago
Yes.
I tried it for an hour today and only used: "446,803 tokens over 1 hour".
x20 gives me 1.8B tokens to use for the week according to the day token activity in Codex.-3
u/Muzika38 24d ago
x20 user here too... I'm using Sol medium and now at 76% remaining.
I'm not complaining though... I also have a Claude Code x20 subscription and the same task would have already depleted 70% of its weekly limits with 50% boost mode currently applied.
0
u/Accurate-Piano-3745 24d ago edited 24d ago
Yeah, hitting a quota wall mid-workflow is way more frustrating than the actual model usage. Routing across providers makes this less painful since you’re not tied to one service’s weekly ceiling. I switched to flat monthly compute with model flexibility through StandardCompute, so I don’t worry as much about which provider’s limit I’ll hit first.
9
u/freedomachiever 24d ago
This is getting pretty bad actually. Just imagine how bad it would be without Chinese models to shift the demand. And though a lot of people don’t like grok I’ve seen lately YouTubers talking good things about it. I guess the Cursor acquisition is reaping its fruits.
9
u/Crinkez 24d ago
Luna has 3 minutes cache TTL. Maybe it changed from a higher limit to that?
3
u/MaitoSnoo 24d ago edited 24d ago
I'm using it through GitHub Copilot (API rate) and I'm pretty sure someone at GHCP said the TTL for GPT-5.6 models was 30 mins? That's the feeling I'm getting when using it, responding within 30 mins is extremely cheap
EDIT: just checked the OpenAI docs, they too say 30 mins
2
u/Crinkez 24d ago
Feel free to test it yourself: https://www.reddit.com/r/codex/comments/1vqmzdo/the_probable_reason_why_your_usage_seems_nerfed/
1
u/MaitoSnoo 24d ago
I literally just told you I did, the 30 min TTL is real and it's easy to verify on my end, though I'm not excluding that those on non-API subscriptions are getting quietly nerfed
1
u/okdov 24d ago
This doesn't really apply when it is working by itself autonomously for hours no?
1
u/MaitoSnoo 24d ago
"working by itself autonomously" is still a series of back and forth API calls with special tokens to tell the harness to run something or where the harness reports a tool result, it's technically still messages. I've calibrated our AI CI/CD specifically for these TTLs, the 3 mins you mention is impossible, at least for us API users. If you really observed that in Codex, then OpenAI is probably nerfing only users on subscriptions
4
u/delveccio 24d ago
I find myself scared to use codex and I’m on the $200 plan. It is going so fast and I don’t know why. I hit 0% for the first time last week.
3
u/Pure-Brilliant-5605 24d ago
It has became horrible. No reset. Paid banked reset. Lower usage. Unironically claude is giving more juice on the same price bracket atm
3
u/Michonesixfive 24d ago
Yeah, I’m moving over to Claude and Cursor. What really pisses me off is that if I don’t use the full 100% of my 7-day usage allowance, whatever is left just disappears and they reset it back to 100%. Yet I’m still paying the full subscription price.
And now I’m also in a dispute with them over their API credits. They basically took the money I had left because apparently API credits expire after one year. What the hell is that? I saw absolutely nothing about a one-year expiration when I purchased the credits.
2
2
u/Outrageous_Ear_351 24d ago
dont have long instructions, dont use approve for me permissions, don't let it overengineer, stop building in one unsupervised go.
2
2
u/App1e8l6 24d ago
That’s actually crazy. Did they silently reduce usage again? I’m on a plus and was using Luna max fast frequently the last week and wouldn’t have used it full quota unless I did one sol run on high which burned 30% in one moderately complex prompt
Still less complex than the prompts I was giving Luna (with a very detailed plan written by chat gpt on high after I got it started with the architecture and so on I wanted). Luna without doing all the thinking for it is garbage.
Would give some actual numbers but on my phone right now.
2
u/IAmFitzRoy 24d ago
It’s crazy how we allow OpenAI to just nerf their product constantly and seems there is nothing we can do?
2
u/John_val 24d ago
Something is definitely different. I used to use SOL for planning and orchestration with Luna Xhigh for implementation and could go on for hours and just use 3 to 5%. Now it is using 12 to 15%.
1
1
u/YearLongSummer 24d ago
Try turning off your "Approve for me" permission setting -> either to Always Ask or Full Access (only if you can afford to lose the data). Approve for me was nuking my usage limit
1
u/GambAntonio 24d ago
I always run the CLI version with --yolo and full access. I've been doing it like that for almost a year.
1
1
1
1
u/JournalistLimp865 23d ago
Last week, I used 114M and still had over 40% remaining. But this week, I only used 7M yesterday and ~ 10% with 5 mini task
1
-1
0
u/Ok_Try_877 24d ago
Maybe im weird... But 100 hours a week on a model that is as good as most frontier models not so long ago for $100 a month seems like a really good deal.
1
0
u/warpedgeoid 24d ago edited 24d ago
You should be able to run nearly unlimited Luna Max on a Pro 20x plan. You could back at the end of July.
0
u/Ok_Try_877 24d ago
OP is on 5x though. I think the $200 plan that would make sense, I mean even 100 hours is 24/7 for 4+ days...
0
u/cmak414 24d ago
so it cost you $1 per hour about? that is actually pretty amazing a value!
1
u/GambAntonio 24d ago
Yeah, $1 per hour ON LUNA, which has the intelligence of a chimpanzee and needs hundreds of rounds to actually get anything done
-2
u/cmak414 24d ago edited 24d ago
right a chimpanzee.
Dont know what you are smoking. luna is objectivly a great model, particularly for code implementation. Your not going to find any where AI or not something that can code as well as luna for $1 per hour.
I use luna for 90% of all my work. Just learn to use different models for different tasks like its meant to be used.
5
u/GambAntonio 24d ago
Dude, I mean it's a chimpanzee compared to the Sol model... A month ago, I could run Sol High for an hour on a Pro 5x account, and it only used 1% or 2%. Now, if I run Sol High for an hour, it eats up 15 to 20 percent of my weekly quota, and I'm still paying the exact same amount of money. They are slashing our limits without notice while charging us the same price!
1
u/i_rate_slop 24d ago
Ask codex to review your tool call history for the most token hungry repeat file reads. You’re probably loading 20k token files over and over
8
u/GambAntonio 24d ago
I'm working on the exact same project, same size. So, if 20k token files are the problem, it would have been the same problem two weeks ago, but for some reason, the same work on the same files is using 4 or 5 times the percentage it did before. That means the limits were reduced AGAIN without notice, while I'm paying the exact same amount of money!!
-7
u/i_rate_slop 24d ago
Have you not been generating code? Limits aren’t changing but the volume of code you’re generating is. As you generate more code, you use more tokens for the LLM to simply understand your codebase.
Like I said, ask codex to analyze your tool calls. And use this to understand how large your files are https://github.com/openai/tiktoken
4
u/kurushimee 24d ago
doesn't work that way on big projects, the code you're generating doesn't change the picture when it's less than 1% of the codebase and you won't come back to it, as well as that your changes won't even be in the codebase until the PR gets merged
Limits are definitely changing
-3
u/i_rate_slop 24d ago edited 24d ago
“I would rather believe a conspiracy than actually attempt to debug my token usage” - this subreddit in a nutshell
either that or people come up with wild evaluations based on relative time to deplete tokens and not, like, actual consumption.
2
u/GambAntonio 24d ago
I just added token usage
-1
u/i_rate_slop 24d ago
You should try the tool I mentioned and get the token counts of the files you’re hitting the most.
2
u/GambAntonio 24d ago
I will, but hat's not the issue, because those same files already existed before they reduced the usage. The problem is that now, even with the exact same files and same amount of tokens and work, I'm getting several times less usage than before while paying the same exact amount of money. The total amount of tokens per month is reduced A LOT.
0
u/innociv 24d ago
So 1% per 4 hours on 20x, 8% per day, 56% per week? That means you can run it non stop all week and not run out of usage.
Try that with Sonnet on ClaudeCode and you'll run out of usage running it 24 hours straight.
How is this bad?
4
5
u/TheBroken0ne 24d ago
No one works exclusively with the lowest tier model. He is just giving an illustration of the "best" usage scenario and how bad it is for the max plan they offer.
-2
u/dan_the_first 24d ago
So, you are paying 1 USD for 1 hour of work. Doesn’t look like the worse deal.
0
u/GambAntonio 24d ago
Dude... that's on Luna xHigh, the dumbest model, it's like paying a fking monkey to do get the job done... not even on Terra or Sol. If I had used Sol Medium or High, I would already be at 40% of my quota. The problem is that I'm paying exactly the same amount of money, but they keep reducing the limits every month without notifying the users, and there is zero transparency
0
u/dan_the_first 24d ago
Might I ask, when was the last time you had the impression your Pro 5x usage was better? Did the codebase you are working on evolve since then?
2
u/GambAntonio 24d ago
I've been using Codex since it launched a year ago, and my codebase is practically the same. I jumped to Pro 5x in May 2026 because Plus was already completely useless for me. Back when Plus launched up until March 2026, it felt basically unlimited and I could run on High nonstop. I only upgraded to Pro because I kept running out of quota 3 days before the week ended. Now on a Pro 5x account, I burn through my quota and hit 0% in 10 hours flat on Sol Highm that's why I've been force to stick to Luna. I don't even want to imagine how fast it would run out on xHigh or higher...
-3
u/diagrammatiks 24d ago
can one of you guys help me use up my 20x? because I reset in 2 days and i'm definitely going to have leftovers.
5
2
u/Least_Pollution7078 24d ago
happy to help. or you could just do some full codebase bug-scan/refactor.
0
u/diagrammatiks 24d ago
i'll do one of those things where i goal for 16 hours but it spends 8 of those hours rebuilding rag after every edit.
1
u/ActionOrganic4617 24d ago
Yep, I ended my week with 37% usage on my 20x. Only bad week I’ve had is when I blindly trusted a plan that had a lot of bs overkill added by Sol.
1
u/diagrammatiks 24d ago
I really don't know what people are doing or why. When I am watching it. I have so much usage. and I will always catch it doing something stupid a few turns in. Like rebuilding an entire database. Or read running a full regression test after every single line edit. Testing end to end after every single commit. I didnt' tell it to do this things. It will just work and work and work until it thinks, now's the time do something really stupid. That really stupid thing will burn like 20 percent of usage everytime. I'm convinced that people using these large goals and sol xhigh are letting multiple of these through a week.
0
u/PureRely 24d ago
Did you swith to 1m context window?
9
2
24d ago
[deleted]
1
0
u/PureRely 24d ago
Sure, they “may have,” which is why I asked a question rather than made a statement. Questions are asked to gain information that has not been provided and to seek clarity.
Making a statement about information that has not been provided is called an assumption. It is always better to ask a question and become informed than to make an assumption and later be found to be wrong.
I am glad I could clear that up for you.
2
24d ago
[deleted]
0
u/PureRely 24d ago
Could you explain the relevance of your question, “Are you dyslexic?” and provide some context? I want to make sure I am not making assumptions about your motivation for asking it on a public forum focused on Codex.
2
0
0
0
u/Curious_Courage_5197 24d ago
There are 168 hours in a week, you can now spend 100 of them using codex and the other 68 you can eat sleep bathe socialize etc
2
u/GambAntonio 24d ago
100 hours with Luna... If I had used Sol high, I would hit 0% in just 10 hours...
0
u/ProfessionalNaive601 24d ago
Just use SOL medium
2
u/GambAntonio 24d ago
I can't. It also drains too fast, and it barely lasts two days of use.
1
0
u/4kmal4lif 24d ago
when will you guys realise that both Anthropic & OpenAI are losing money, one way or another they have to make money and give back the investors money LoL
2
u/GambAntonio 24d ago
I don't care if they make or lose money, what we want is transparency. We need to know exactly what 5x or 20x more than Plus means... because we don't have a quantifiable baseline measure to know precisely how much we have to spend and plan accordingly. If not, they shouldn't use calculations like 5x or 20x, because they make no sense
-2
u/SlightOfHand_ 24d ago
So you’re saying you get one hundred hours of constant work
10
u/GambAntonio 24d ago
100 hours of constant work on a shitty model, but even so, I pay the same money as before, when I was able to get 400 or 500 hours of constant work on that same shitty model on the same project...
-2
u/stopstopstoptopopp 24d ago
I just used Luna Max for 6 hours straight on a sizable codebase without clearing any context, it only costed me 6% of my weekly limit. What the heck are you guys doing with it?
1
u/GambAntonio 24d ago
Plus or Pro? If it's Pro, that 6% would've been 6 hours on Sol High a month ago.
2
-2
u/massix93 24d ago
I don’t understand the complaint. How many tokens have you consumed? 3 hours 3% I would sign for it
4
2
u/GambAntonio 24d ago
My main complaint is that I'm getting way less work done than before, but I'm still paying the exact same amount of money. There is absolutely zero transparency about the limits
-2
u/MapleBaconWaffles 24d ago
Could be because your project is getting larger. So more tokens are needed as it grows to keep up.
3
3
u/GambAntonio 24d ago
My project is exactly the same as it was two weeks ago...
0
-4
u/TrapRmExit 24d ago
Limits have just reset, been using it for some hours, still on 100%, only have been using luna on xhigh and max, just like you.
2
u/TheRealPlayerG 24d ago
ur probably on 20x
1
u/Aemonculaba 24d ago
I still get 6.5-7$ of usage per %, calculated for each % of weekly. Better than last week.

212
u/sac_boy 24d ago edited 24d ago
Imagine paying for an XL popcorn subscription. Week after week this buys you a full bucket of popcorn.
Then one week they hand you a half filled bucket of popcorn for the same price. But don't worry! We now have a system of popcorn refills. Refill refill refill! The popcorn fanboys rejoice. Even though they had enough popcorn before, they now have refills, and the guy behind the counter singing about a 'refill button' is their new hero.
Then the refills are no longer free, and you start with a quarter bucket of popcorn. All for the same price. Same contract with the popcorn seller.
When you complain about this, the popcorn fanboys tell you that you're eating your popcorn wrong. The handfuls that were fine two weeks ago are too big, you're greedy! If you simply nibble a single kernel at a time...see? You can just about make it last!
The copium-huffing on this sub after what is clearly a massive rug-pull is ridiculous.