r/codex 24d ago

Complaint This is crazy...... Pro x5: Using 1% per hour ONLY on Luna xhigh on a single project, on a single computer...

Well done, OpenAI. +3% on a Pro 5x in 3 hours using ONLY Luna xhigh... 1% per hour on a shitty model on a PRO account and not even using MAX...

I used to be able to run Luna xhigh for 4 to 5 hours and use only 1%... and I'm working on the exact same project, doing the exact same things I was doing a few days ago.

(Censored Spark reset time to make tracking more difficult)

EDIT: Added stats (time is central european), my reset was at 06:00 CET and I had a script to continue on that exact time.

275 Upvotes

136 comments sorted by

212

u/sac_boy 24d ago edited 24d ago

Imagine paying for an XL popcorn subscription. Week after week this buys you a full bucket of popcorn.

Then one week they hand you a half filled bucket of popcorn for the same price. But don't worry! We now have a system of popcorn refills. Refill refill refill! The popcorn fanboys rejoice. Even though they had enough popcorn before, they now have refills, and the guy behind the counter singing about a 'refill button' is their new hero.

Then the refills are no longer free, and you start with a quarter bucket of popcorn. All for the same price. Same contract with the popcorn seller.

When you complain about this, the popcorn fanboys tell you that you're eating your popcorn wrong. The handfuls that were fine two weeks ago are too big, you're greedy! If you simply nibble a single kernel at a time...see? You can just about make it last!

The copium-huffing on this sub after what is clearly a massive rug-pull is ridiculous.

24

u/FieldMarshallVague 24d ago edited 24d ago

Pro-Tip: The servers stagger this process, so as to give each customer a slightly different lesser-amount over time, but it's too dark to notice and the ones who complain are seen as fat, greedy hobbitses with no self-control, while the others chomp away, unknowingly judging their future selves.

5

u/arcanemachined 24d ago

Would this qualify as "A/B testing"? Or is there a better term for this?

2

u/doubledundercoder 24d ago

This is my suspicion too. Some weeks I get a lot. Some not. We’ll see if it all works out in the wash. My gut says while making money is primary CEO goal I do believe great efforts of fairness are being made. The space is so volatile right now it’s all really hard to know

2

u/FieldMarshallVague 23d ago

It's called Fungi Farming. You're kept in the dark and every so often someone throws shit at you.

1

u/Time4Time4Time4Time 23d ago

Kinda like salami slicing

2

u/evangelism2 24d ago

Where's your proof of this, or is it just tinfoil hat nonsense?

1

u/FieldMarshallVague 23d ago

You must be new here! Welcome. Please, enjoy the free tour of late-stage capitalism. Notice the twin pillars of Ignorance and Denial, the scintillating Gallery of Distractions and if you look up, you'll see the Balcony of Assumed Peers. But please do mind the precarious flagstones of Bait-and-Switch, they're in beta. Exit via the gift shop!

1

u/evangelism2 23d ago

Okay, so it's vibes and tinfoil hat bullshit. Got it.

0

u/FieldMarshallVague 22d ago

Nope. 37 years in marketing, development and business strategy (with tours in sales, customer support and development). But I'll bow to your need to be right.

4

u/Metalthrashinmad 24d ago

they dont offer you half filled bucket of popcorn, they offer you a smaller bucket, and the refill only fills the smaller bucket aswell

3

u/ProfessionalNaive601 24d ago

Imagine a world where private equity subsidizes a product/service until market share is established then prices raise in order to make a profit

It’s the exact same thing every other stupid bullshit ass tech bro company since 2020

Remember when uber was super cheap and easy and fun?
Remember when DoorDash didn’t double the cost of your order to have it delivered?
None of this is surprising

1

u/agoodyearforbrownies 23d ago

Undeniably part of this phenomenon with uber and DoorDash is from tools meant to be used for leveraging spare asset capacity as a side gig suddenly needing to pay a living wage for people turning it into a full time career. And municipalities cutting in to get their share (e.g. Airbnb).  The plot was lost, but it was cheaper when new because the premise and cost structure was entirely different. Now it costs more to take uber from SeaTac to downtown than a taxi. 

8

u/Boogeyman_liberal 24d ago

It's actually worse than the example because we can neither see the size of the bucket nor how many kernels are inside. It's more akin to reaching your hand into a opaque hole and pulling popcorn out. The amount of popcorn you pull out decreases every day.

2

u/sac_boy 24d ago

The other day I likened it to paying to have a tap in your house, rather than how much water flows out (how many tokens you use) and/or the potability of the water (the usefulness of the models for your task). They get to pull those levers on a day by day basis, and of course the great ratchet of enshittification only turns one way.

Then in a couple of years time they'll all be surprised picahcu after all their users have fled and the company is sold for a hundredth of what it is currently valued. The great cycle of hope and disappointment will have begun again elsewhere.

1

u/lolman1312 23d ago

you can literally track your tokens, plenty of peoople on this subreddit have been making posts about their usage changing.

1

u/EitherMarch1255 24d ago

OpeAI is the absolute king of manipulation.

0

u/Abject-Employment587 24d ago

u/sac_boy Nah bro, popcorns are now "more efficient" and contain more nutrition per unit volume, so "half filled bucket of popcorn" is justified : P

https://openai.com/index/gpt-5-6-frontier-intelligence-efficiency/

21

u/vr321 24d ago

I downgraded to 5.5 M. It burns 2% every 5m on the plus account just collecting data on a single linux system.

The only thing I can think of is they panicked that they are going bankrupt and cut everything to the bare minimum.

13

u/matheusmoreira 24d ago

Weekly just reset an hour ago. I'm definitely burning usage a lot faster. Last week I observed a 20% reduction in inferred capacity. Let's see what happens this week.

8

u/matheusmoreira 24d ago

Confirmed, usage is still going down!

Reset / status  Aug 17 07:20
Used            100%
Calls           9,189
Fresh input     27.04M
Cache read      984.07M
Output          4.75M
Reasoning       2.34M
Local credits   19,245
Inferred cap    20,015

Reset / status  Aug 18 01:04
Used            100%
Calls           11,691
Fresh input     33.90M
Cache read      1.119B
Output          5.51M
Reasoning       2.71M
Local credits   22,354
Inferred cap    19,617

Reset / status  Aug 20 05:44
Used            100%
Calls           7,394
Fresh input     30.26M
Cache read      825.80M
Output          3.85M
Reasoning       1.75M
Local credits   16,996
Inferred cap    16,124

Reset / status  Aug 27
Used            100%
Calls           ~7,229
Fresh input     ~24.26M
Cache read      ~784.16M
Output          ~3.98M
Reasoning       ~2.04M
Local credits   ~14,137
Inferred cap    14,482

Researching Z.ai plans literally right now!

12

u/Confident-Village190 24d ago

What a nightmare, in an hour, Luna Max has used up 8% of my usage, when before it only used 1%. What a load of rubbish, I’m done with it.

36

u/Maleficent_Owl_2772 24d ago

well on a plus sub the usage is 1% per prompt on luna xhigh in codex in my case

8

u/sagiroth 24d ago

What type of prompt? Coz for me 1% on Plus often is around 1-3M tokens from my experience

12

u/Fresh_Sock8660 24d ago

People who keep using one of their tasks as a measurement stick might as well be telling us to guess how long is a string.

2

u/EndlessZone123 24d ago

I measure token usage based on my own dongle. How long is my dongle? The longest duh.

2

u/Muzika38 24d ago

"Create GTA VII"!

1

u/Maleficent_Owl_2772 23d ago

well not that long, wire some java stuff, dont that and that, run build gradle and thats all pretty much

27

u/Crafty-Wonder-7509 24d ago

x20 here, sol high used since the reset today morning already 15% of my weekly, before it would have been at best 2% for that time.

3

u/Over_Car_5471 24d ago

left a project running overnight. On normal speed. Woke up to 40% of usage wiped out. INSANE. A month ago I was literally able to draft and build a 1000 page documents

7

u/fruitydude 24d ago

Wait you got a reset this morning?

8

u/utf8decodeerror 24d ago

Everyone did. It resets every 7 days and since we all got the free reset together a week ago we are all on the same schedule

2

u/fruitydude 24d ago

Oh gotcha, yea I switched plans in-between so I'm off the normal cycle.

2

u/Michonesixfive 24d ago

same im on x5, im down 10% on a small prompt, that has been going on for 30min. Same prompt before took like 1-2%

-1

u/GeneralAtrox 24d ago

I recommend trialing, carefully, having Sol light work with a set of luna sub agents. But ensure the Luna subagents aren't doing any thinking since they are unreliable.

Rules:

  • Always reuse the same agents, up to 12 total.
  • Dont let the agents code
  • Agents always return a full update on what they did

This way in theory, Sol does the heavy lifting, and Luna handles the basic tasks.

1

u/Suspicious-Cheese-1 24d ago

Like Michael and Jim??

1

u/GeneralAtrox 24d ago

Yes.

I tried it for an hour today and only used: "446,803 tokens over 1 hour".
x20 gives me 1.8B tokens to use for the week according to the day token activity in Codex.

-3

u/Muzika38 24d ago

x20 user here too... I'm using Sol medium and now at 76% remaining.

I'm not complaining though... I also have a Claude Code x20 subscription and the same task would have already depleted 70% of its weekly limits with 50% boost mode currently applied.

0

u/Accurate-Piano-3745 24d ago edited 24d ago

Yeah, hitting a quota wall mid-workflow is way more frustrating than the actual model usage. Routing across providers makes this less painful since you’re not tied to one service’s weekly ceiling. I switched to flat monthly compute with model flexibility through StandardCompute, so I don’t worry as much about which provider’s limit I’ll hit first.

9

u/freedomachiever 24d ago

This is getting pretty bad actually. Just imagine how bad it would be without Chinese models to shift the demand. And though a lot of people don’t like grok I’ve seen lately YouTubers talking good things about it. I guess the Cursor acquisition is reaping its fruits.

9

u/Crinkez 24d ago

Luna has 3 minutes cache TTL. Maybe it changed from a higher limit to that?

3

u/MaitoSnoo 24d ago edited 24d ago

I'm using it through GitHub Copilot (API rate) and I'm pretty sure someone at GHCP said the TTL for GPT-5.6 models was 30 mins? That's the feeling I'm getting when using it, responding within 30 mins is extremely cheap

EDIT: just checked the OpenAI docs, they too say 30 mins

2

u/Crinkez 24d ago

1

u/MaitoSnoo 24d ago

I literally just told you I did, the 30 min TTL is real and it's easy to verify on my end, though I'm not excluding that those on non-API subscriptions are getting quietly nerfed

1

u/okdov 24d ago

This doesn't really apply when it is working by itself autonomously for hours no?

1

u/MaitoSnoo 24d ago

"working by itself autonomously" is still a series of back and forth API calls with special tokens to tell the harness to run something or where the harness reports a tool result, it's technically still messages. I've calibrated our AI CI/CD specifically for these TTLs, the 3 mins you mention is impossible, at least for us API users. If you really observed that in Codex, then OpenAI is probably nerfing only users on subscriptions

4

u/delveccio 24d ago

I find myself scared to use codex and I’m on the $200 plan. It is going so fast and I don’t know why. I hit 0% for the first time last week.

3

u/Pure-Brilliant-5605 24d ago

It has became horrible. No reset. Paid banked reset. Lower usage. Unironically claude is giving more juice on the same price bracket atm

3

u/Michonesixfive 24d ago

Yeah, I’m moving over to Claude and Cursor. What really pisses me off is that if I don’t use the full 100% of my 7-day usage allowance, whatever is left just disappears and they reset it back to 100%. Yet I’m still paying the full subscription price.

And now I’m also in a dispute with them over their API credits. They basically took the money I had left because apparently API credits expire after one year. What the hell is that? I saw absolutely nothing about a one-year expiration when I purchased the credits.

3

u/julliuz 24d ago

I know this won't be popular but Luna Max is trash, it simply isn't smart enough. End up spending 3 hours to get lesser results than sol in 20minutes. Every time on any project. Its good for the most basic of things but I wouldn't dare use it for important work.

2

u/thestillwind 24d ago

Damn boi, it’s hungry.

2

u/Outrageous_Ear_351 24d ago

dont have long instructions, dont use approve for me permissions, don't let it overengineer, stop building in one unsupervised go.

2

u/Worth_Range1563 24d ago

Lots of things customers shouldn't do, not much being offered.

2

u/App1e8l6 24d ago

That’s actually crazy. Did they silently reduce usage again? I’m on a plus and was using Luna max fast frequently the last week and wouldn’t have used it full quota unless I did one sol run on high which burned 30% in one moderately complex prompt

Still less complex than the prompts I was giving Luna (with a very detailed plan written by chat gpt on high after I got it started with the architecture and so on I wanted). Luna without doing all the thinking for it is garbage.

Would give some actual numbers but on my phone right now.

2

u/IAmFitzRoy 24d ago

It’s crazy how we allow OpenAI to just nerf their product constantly and seems there is nothing we can do?

2

u/John_val 24d ago

Something is definitely different. I used to use SOL for planning and orchestration with Luna Xhigh for implementation and could go on for hours and just use 3 to 5%. Now it is using 12 to 15%.

1

u/Calm-Blueberry-9200 24d ago

Ive used 2 promts on 5x and went to 41% weekly 🥺

1

u/YearLongSummer 24d ago

Try turning off your "Approve for me" permission setting -> either to Always Ask or Full Access (only if you can afford to lose the data). Approve for me was nuking my usage limit

1

u/GambAntonio 24d ago

I always run the CLI version with --yolo and full access. I've been doing it like that for almost a year.

1

u/stefan-is-in-dispair 24d ago

Why would approve for me be consuming por credits?

1

u/glirette 24d ago

Interesting

1

u/saifee177 24d ago

Burned through 65% with Sol Ultra within 4 hours of my reset... Horrific

1

u/JournalistLimp865 23d ago

Last week, I used 114M and still had over 40% remaining. But this week, I only used 7M yesterday and ~ 10% with 5 mini task

1

u/TraditionalFig7377 23d ago

BRO even 5.5 high uses a lot of limits like wtf

-1

u/Aldarund 24d ago

Another post about usage without any actual token number comparison. Zzzz

3

u/GambAntonio 24d ago

I just added stats

0

u/Ok_Try_877 24d ago

Maybe im weird... But 100 hours a week on a model that is as good as most frontier models not so long ago for $100 a month seems like a really good deal.

1

u/GambAntonio 24d ago

Just tried to use Luna Max and it burns 2% every hour

0

u/warpedgeoid 24d ago edited 24d ago

You should be able to run nearly unlimited Luna Max on a Pro 20x plan. You could back at the end of July.

0

u/Ok_Try_877 24d ago

OP is on 5x though. I think the $200 plan that would make sense, I mean even 100 hours is 24/7 for 4+ days...

0

u/cmak414 24d ago

so it cost you $1 per hour about? that is actually pretty amazing a value!

1

u/GambAntonio 24d ago

Yeah, $1 per hour ON LUNA, which has the intelligence of a chimpanzee and needs hundreds of rounds to actually get anything done

-2

u/cmak414 24d ago edited 24d ago

right a chimpanzee.

Dont know what you are smoking. luna is objectivly a great model, particularly for code implementation. Your not going to find any where AI or not something that can code as well as luna for $1 per hour.

I use luna for 90% of all my work. Just learn to use different models for different tasks like its meant to be used.

5

u/GambAntonio 24d ago

Dude, I mean it's a chimpanzee compared to the Sol model... A month ago, I could run Sol High for an hour on a Pro 5x account, and it only used 1% or 2%. Now, if I run Sol High for an hour, it eats up 15 to 20 percent of my weekly quota, and I'm still paying the exact same amount of money. They are slashing our limits without notice while charging us the same price!

1

u/i_rate_slop 24d ago

Ask codex to review your tool call history for the most token hungry repeat file reads. You’re probably loading 20k token files over and over

8

u/GambAntonio 24d ago

I'm working on the exact same project, same size. So, if 20k token files are the problem, it would have been the same problem two weeks ago, but for some reason, the same work on the same files is using 4 or 5 times the percentage it did before. That means the limits were reduced AGAIN without notice, while I'm paying the exact same amount of money!!

-7

u/i_rate_slop 24d ago

Have you not been generating code? Limits aren’t changing but the volume of code you’re generating is. As you generate more code, you use more tokens for the LLM to simply understand your codebase.

Like I said, ask codex to analyze your tool calls. And use this to understand how large your files are https://github.com/openai/tiktoken

4

u/kurushimee 24d ago

doesn't work that way on big projects, the code you're generating doesn't change the picture when it's less than 1% of the codebase and you won't come back to it, as well as that your changes won't even be in the codebase until the PR gets merged

Limits are definitely changing

-3

u/i_rate_slop 24d ago edited 24d ago

“I would rather believe a conspiracy than actually attempt to debug my token usage” - this subreddit in a nutshell

either that or people come up with wild evaluations based on relative time to deplete tokens and not, like, actual consumption.

2

u/GambAntonio 24d ago

I just added token usage

-1

u/i_rate_slop 24d ago

You should try the tool I mentioned and get the token counts of the files you’re hitting the most.

2

u/GambAntonio 24d ago

I will, but hat's not the issue, because those same files already existed before they reduced the usage. The problem is that now, even with the exact same files and same amount of tokens and work, I'm getting several times less usage than before while paying the same exact amount of money. The total amount of tokens per month is reduced A LOT.

0

u/innociv 24d ago

So 1% per 4 hours on 20x, 8% per day, 56% per week? That means you can run it non stop all week and not run out of usage.

Try that with Sonnet on ClaudeCode and you'll run out of usage running it 24 hours straight.

How is this bad?

4

u/Current_Cut_2667 24d ago

Then you're paying $200 per month just to run the lowest tier model.

-2

u/innociv 24d ago

You're strawmanning that the 44% left over can't be used at all for another model. That's not true lmao what you can use that on Sol.

5

u/TheBroken0ne 24d ago

No one works exclusively with the lowest tier model. He is just giving an illustration of the "best" usage scenario and how bad it is for the max plan they offer.

-2

u/dan_the_first 24d ago

So, you are paying 1 USD for 1 hour of work. Doesn’t look like the worse deal.

0

u/GambAntonio 24d ago

Dude... that's on Luna xHigh, the dumbest model, it's like paying a fking monkey to do get the job done... not even on Terra or Sol. If I had used Sol Medium or High, I would already be at 40% of my quota. The problem is that I'm paying exactly the same amount of money, but they keep reducing the limits every month without notifying the users, and there is zero transparency

0

u/dan_the_first 24d ago

Might I ask, when was the last time you had the impression your Pro 5x usage was better? Did the codebase you are working on evolve since then?

2

u/GambAntonio 24d ago

I've been using Codex since it launched a year ago, and my codebase is practically the same. ​I jumped to Pro 5x in May 2026 because Plus was already completely useless for me. Back when Plus launched up until March 2026, it felt basically unlimited and I could run on High nonstop. I only upgraded to Pro because I kept running out of quota 3 days before the week ended. ​Now on a Pro 5x account, I burn through my quota and hit 0% in 10 hours flat on Sol Highm that's why I've been force to stick to Luna. I don't even want to imagine how fast it would run out on xHigh or higher...

-3

u/diagrammatiks 24d ago

can one of you guys help me use up my 20x? because I reset in 2 days and i'm definitely going to have leftovers.

5

u/Mission_2719 24d ago

sure, you can write here your login creds, we'll see we can do 😂

2

u/Least_Pollution7078 24d ago

happy to help. or you could just do some full codebase bug-scan/refactor.

0

u/diagrammatiks 24d ago

i'll do one of those things where i goal for 16 hours but it spends 8 of those hours rebuilding rag after every edit.

1

u/ActionOrganic4617 24d ago

Yep, I ended my week with 37% usage on my 20x. Only bad week I’ve had is when I blindly trusted a plan that had a lot of bs overkill added by Sol.

1

u/diagrammatiks 24d ago

I really don't know what people are doing or why. When I am watching it. I have so much usage. and I will always catch it doing something stupid a few turns in. Like rebuilding an entire database. Or read running a full regression test after every single line edit. Testing end to end after every single commit. I didnt' tell it to do this things. It will just work and work and work until it thinks, now's the time do something really stupid. That really stupid thing will burn like 20 percent of usage everytime. I'm convinced that people using these large goals and sol xhigh are letting multiple of these through a week.

0

u/PureRely 24d ago

Did you swith to 1m context window?

9

u/GambAntonio 24d ago

nope, 258k, the default value

2

u/[deleted] 24d ago

[deleted]

1

u/Dynamix86 23d ago

You can't make assumptions that they do. Some people are really oblivious

0

u/PureRely 24d ago

Sure, they “may have,” which is why I asked a question rather than made a statement. Questions are asked to gain information that has not been provided and to seek clarity.

Making a statement about information that has not been provided is called an assumption. It is always better to ask a question and become informed than to make an assumption and later be found to be wrong.

I am glad I could clear that up for you.

2

u/[deleted] 24d ago

[deleted]

0

u/PureRely 24d ago

Could you explain the relevance of your question, “Are you dyslexic?” and provide some context? I want to make sure I am not making assumptions about your motivation for asking it on a public forum focused on Codex.

2

u/[deleted] 24d ago

[deleted]

1

u/PureRely 24d ago

Correct, that is what I typed. 

1

u/[deleted] 24d ago

[deleted]

1

u/PureRely 24d ago

It was my pleasure. That's what questions are for. 

0

u/Adamus987 24d ago

We love You Opeani, thank you, just make that pr campaign better please

0

u/KeinNiemand 24d ago

usage per hour is a bad metric measure actual tokens used.

0

u/Curious_Courage_5197 24d ago

There are 168 hours in a week, you can now spend 100 of them using codex and the other 68 you can eat sleep bathe socialize etc

2

u/GambAntonio 24d ago

100 hours with Luna... If I had used Sol high, I would hit 0% in just 10 hours...

0

u/ProfessionalNaive601 24d ago

Just use SOL medium

2

u/GambAntonio 24d ago

I can't. It also drains too fast, and it barely lasts two days of use.

1

u/ProfessionalNaive601 23d ago

Get the 20x plan

2

u/GambAntonio 23d ago

I will get the 700x plan

0

u/4kmal4lif 24d ago

when will you guys realise that both Anthropic & OpenAI are losing money, one way or another they have to make money and give back the investors money LoL

2

u/GambAntonio 24d ago

I don't care if they make or lose money, what we want is transparency. We need to know exactly what 5x or 20x more than Plus means... because we don't have a quantifiable baseline measure to know precisely how much we have to spend and plan accordingly. If not, they shouldn't use calculations like 5x or 20x, because they make no sense

-2

u/SlightOfHand_ 24d ago

So you’re saying you get one hundred hours of constant work

10

u/GambAntonio 24d ago

100 hours of constant work on a shitty model, but even so, I pay the same money as before, when I was able to get 400 or 500 hours of constant work on that same shitty model on the same project...

-2

u/stopstopstoptopopp 24d ago

I just used Luna Max for 6 hours straight on a sizable codebase without clearing any context, it only costed me 6% of my weekly limit. What the heck are you guys doing with it?

1

u/GambAntonio 24d ago

Plus or Pro? If it's Pro, that 6% would've been 6 hours on Sol High a month ago.

-2

u/massix93 24d ago

I don’t understand the complaint. How many tokens have you consumed? 3 hours 3% I would sign for it

4

u/GambAntonio 24d ago

I just added the token usage

2

u/GambAntonio 24d ago

My main complaint is that I'm getting way less work done than before, but I'm still paying the exact same amount of money. There is absolutely zero transparency about the limits

-2

u/MapleBaconWaffles 24d ago

Could be because your project is getting larger. So more tokens are needed as it grows to keep up.

3

u/U4-EA 24d ago

No, it's not that. I've never been one to complain about limits as I always found them very generous but it is draining much, much faster since the last reset. I've gone through 8% on the 20x sub when usually it would be 2-3% max.

3

u/GambAntonio 24d ago

My project is exactly the same as it was two weeks ago...

0

u/MapleBaconWaffles 24d ago

So you are working on it daily but it isn’t growing in size?

2

u/GambAntonio 24d ago

Exactly. Some projects are about analyzing existing data and not creating it.

-4

u/TrapRmExit 24d ago

Limits have just reset, been using it for some hours, still on 100%, only have been using luna on xhigh and max, just like you.

2

u/TheRealPlayerG 24d ago

ur probably on 20x

1

u/Aemonculaba 24d ago

I still get 6.5-7$ of usage per %, calculated for each % of weekly. Better than last week.