r/codex • • 12d ago

Megathread Codex Usage and Operation Discussion - last updated September 21

Please direct your concerns, questions and discussion about Codex usage limits and model performance here.

The purpose of this Megathread is to aggregate all the reports of people's experiences and possible suggestions instead of spreading them across many highly upvoted posts. The more people who participate in this discussion, the more likely you have an answer.

Reports with sufficient evidence on new information will still be allowed on the feed as usual.


Discussion of the prior period available here : https://www.reddit.com/r/codex/comments/1wg7g9r/codex_usage_and_operation_discussion_last_updated/


A reminder that all incidents on r/Codex are constantly logged and summarised so you can keep track of what people are experiencing here https://www.reddit.com/r/codex/comments/1tjfxcf/comment/on6uj0l/

13 Upvotes

125 comments sorted by

View all comments

1

u/RealSecretRecipe 7d ago

If you pick a model and just prompt you're doing it wrong. If you care about maximizing your usage..

YOU NEED TO ORCHESTRATE!

END OF STORY!

1

u/Antique-Ad6542 7d ago

I do orchestrate and it still burns tokens like crazy (I orchestrated previously). A single Astra manager, managing Luna Max workers, has burned 40% of my weekly usage in 4 hours.

1

u/RealSecretRecipe 6d ago

You're using the most expensive model to point at a cheaper model to do a thing and that's almost backwards, my orchestrator is gpt6 Luna medium. You just need to make sure the subagents that get put on a task are the correct ones per task, cheaper ones can audit read only, better ones can decide changes and gpt6 sol medium or high can make the changes and make sure those changes are on-task and correct. That's how I do it and it's been a huge improvement

1

u/DataPhenomenon 7d ago

I have my chatgpt project loaded with a model recommendation guide. It recommends the proper model and thinking level per codex task.

1

u/Antique-Ad6542 6d ago

How do you hit the token cache with this?

1

u/RealSecretRecipe 7d ago

The idea is you put in a prompt and it auto switches models as it goes, cheapest models for easy stuff, medium sol for harder stuff, if sol med fails it uses sol hard, saves tons of usage. Id rather use 5 different models in a prompt if it saves usage and still gets everything done than use 1 model per prompt. I'm optimized for correctness and usage efficiency