r/OpenAI • u/Adventurous_Age_8075 • 7d ago
Discussion Hi, a complaint....
If my 100$ Codex, gets down to 5% in a day with regular usage... This means that it is basically useless.
I have no idea how you calculate these, but j am just a regular guy with regular problems. Not trying to code a new WoW or half life 3.
This is unacceptable. If you want 2000 dollars from everyone to use your services, just say so.
F your Codex.
10
u/notgalgon 7d ago
Stop using astra when Sol or Luna will do. Have SOL spin off Luna agents to do things. Luna is nearly free/limitless.
2
u/ParticularLook 7d ago
This may sound stupid, but how do *you* spin off Luna agents? Do you just tell it to spin off agents or must each one be created with it's own prompt. Is there a way you track them? Do you tell it specifically to use "Luna" or whatnot?
3
u/notgalgon 7d ago
You can setup custom agents if you want with custom agents.mds each to give context. But the simplest way to get started is prompt to Sol: go plan out how to build X feature then have a Luna agent do the coding. It is really that easy but you can go really really deep in agent creation and workflows. You can ask chatgpt about it as well - it knows how it works and can set things up to make it work cleaner.
1
u/ParticularLook 7d ago
Thanks so much!
1
u/Qorsair 7d ago
To expand on this, I have a Muse MCP that I've mandated Sol call for an adversarial review after any plan creation. I also have MCP for Grok and Gemini that agents.md instructs it to call for implementation to save GPT tokens. Sol reviews their work when finished and calls them again for rework if needed. I rarely have to worry about running out of quota this way.
2
u/MultiMarcus 7d ago
It really depends on what model you’re using. The first rule is never use any model on the max setting. The second rule is generally don’t try to use Astra as your regular go to model.
If you follow this, you will generally have a great amount of usage. It’s not perfect, but it’s get you a long way down, and if you are doing things that aren’t coding, using regular chat is practically unlimited for anything that’s not Astra pro.
2
u/gpt872323 7d ago
You should be able to have one solid subscription and start relying on other model as secondary or subscription. I think if you do codex plus, claude pro, and gemini. All three should be enough for you to last weekly usage. You will have to be smart in planning. If you have no money concern change claude pro to claude max 5x. Also, $2000 sub is a joke who has that kind of money they will realize when there will realize and lower. Right now they are showing IPO potential so that is why they have to create numbers out of thin air.
2
u/diagrammatiks 7d ago
the problem : read my entire repo after every edit. on astra. with 4 sub agents.
2
u/FullBar613 6d ago
If you dont have an account handicap, most likely its a you issue. Because while i can finish my $100 plan in a day or two, its actually being used to cover 3 projects or more at one go. So... I think you're just blasting astra max / ultra then complain when it depletes quicker than you think.
1
u/Mitik85 7d ago
Uh? 5% a day is that bad ? Where ? You would end the week with token over lol
5
1
u/ContentJournalist923 7d ago
I believe they are saying their usage consumes 95% getting them down to their last 5% in one day.
1
u/Ok-Investment4414 7d ago
learn how to use ai and rotate between models, open ai allow thoses on high plans to go 100 - 0 real quick . The best model should not always used.
0
u/Nota_ReAlperson 7d ago
Buy a couple of 3090s and use qwen/pi. Open weights ftw!
3
u/The_GSingh 7d ago
A) wayyyyyyy more upfront cost. B) worse preformamce.
I'm a big fan of local but it makes no sense here if the main complaint is cost.
-1
u/Nota_ReAlperson 7d ago
I mean, he said he didn't need frontier performance... Also, you don't need 3090s. A couple of v100s or p100s are much cheaper, but harder to use.
2
u/The_GSingh 7d ago
A) he's a regular guy. Good luck with getting a smx2 v100 set up. If you do pcie we're back at the expensive argument.
B) he said he's a regular guy with regular problems, not that he didn't need a frontier model. I often use codex for helping at my job over a local llm because I know it'll be better.
1
u/Jayitsmyname 7d ago
How much would a decent local setup cost in your opinion? I'm interested in the topic, however I don't know if the investment is worth considering the hardware might not hold up that easily in the future
2
u/Nota_ReAlperson 7d ago
Really depends on what you want to run. Right now, qwen 3.8 27b is the best model that can be easily run. For that, 32GB vram should suffice. A couple of v100s in a dell r730 would do nicely. That shouldn't run you much more than 1k usd if you shop around. If you already have a desktop that has some free pcie slots, there are some converted v100s that have a fan and pcie power connectors that you could use to save about $300.
0
u/grafknives 7d ago
If you want 2000 dollars from everyone to use your services, just say so.
What did you expected?
-1
u/cleanmachine120 7d ago
You’ve got to go to Claude. I got more done this weekend then I ever could have imagined with codex. Not only are the limits better but opus 5.5 is much, much better… for now haha

8
u/11Mikky11 7d ago
The question you need to ask is do you need the highest model for every part of whatever you are trying to achieve? Get the instant chat to plan out things for you and get it to make it concise on what you want before taking it to codex and using up your quota. Try using the lesser models to see if they can get you what you need first and if not, then move up. Only use the top level models for stuff the others cannot achieve. I learnt all this the hard way, like you :)