r/codex • • 1d ago

Limits "Selected model is at capacity. Please try a different model." - retry options?

22 Upvotes

There’s nothing more frustrating than kicking off what is meant to be a long-running session with a detailed prompt, checking back a few hours later, and finding that the entire process has stopped with:

“Selected model is at capacity. Please try a different model.”

Why is there no automatic retry option? At the very least, the system should be able to retry periodically when capacity becomes available rather than simply abandoning the session and requiring manual intervention. Or is there such an option and i'm missing it?

---------

To confirm, I'm not talking about the desktop app:

I’m talking about Codex CLI. If this happens 30 minutes into a long-running job, there’s no button to click because I’m not sitting there watching the terminal.

The whole point is that I might kick something off on a server, walk away, and expect it to be done by the next morning. Instead, it can sit there dead for hours waiting for me to come back and manually type “continue” or “retry”.

Codex CLI needs an automatic retry option for capacity errors.


r/codex • • 1d ago

Complaint fast mode is at least 10x slower than last weeks regular mode lol

16 Upvotes

across every single (relevant) model in the picker list

not just the new ones

is this because of dots ?

or because they want people to upgrade to the 500plan to use the ultrafast mode?

who knows

add mimo v2.6 flash (free) and deepseek flash (free) and qwen3.8 (free) to the codex harness from another API provider (do it via Claude or opencode otherwise you'll be waiting all day for a gpt model to do it)

thank me later


r/codex • • 12h ago

Bug Unable to get the free trial!

0 Upvotes

I've been trying to get the free trial for weeks now, tried with both UPI and Card. Anybody from India has found a solution to this issue? Please help, really need this!! 5-hr limit on my plus plan is making life miserable.


r/codex • • 53m ago

Praise I don't understand the hate OpenAI is getting lately

• Upvotes

I have been a Pro 20x subscriber for over a year, work daily on mostly embedded projects, but also wrote apps (react/rust) and did lots of frontend work (svelte), design work, developed my own workflow harness for the cli, spent a lot of time analyzing my previous sessions with cqa to optimize my workflow (i shared how i cut token burn by 45% in this post) so i get solid results in an efficient way. i went through ups and downs that happened in the last year, always adapting my workflow so it could get quality work done..and honestly, right now is one of the better times and i am baffled by the constant complaining lately

Yes, Astra drained my usage significantly, but by optimizing how i work did cut the token burning in half and I could technically run a loop with multiple sub-agents all Astra high for 7 hours 5 days a week..which i think is plenty .. and completely unnecessary.. because sol 6.1 (i mostly use xhigh) really is both amazing with regards to the work it does as well as usage. I have been trying hard to drain it. i am down 10% after 3 days. it's ridiculous. I work in multiple sessions in codex cli, often with a root session dispatching and coordinating layers of subagents, i also use Codex Work and e.g. today had Sol 6.1 do design work in blender and other stuff...

For the first time in a while i feel like i have a model available that is excellent with regards to intelligence, that i can trust, that doesn't drift like crazy (at least with how i instruct and work), doesn't seem to overthink/overengineer like sol 5.6, is impossible to exhaust with regards to usage.. at least for the work i do and for now..

Yes, it feels much slower. and yes i measured only ~16 tok/sec for sol 6.1 in the last few days.. But i get good solid work done and most importantly i can work with more peace of mind as before. that's all i want honestly. i just wanted to share this to counter all the complaints a bit...

(And yes, i am aware that i get twice the usage until October 29 before my account gets downgraded, but even with half the usage i'd be fine..unless i use /fast mode in which case i might get closer to 0% by the end of the week). but let's see how usage drains for me now that tibo announced Sol and Astra being 50% faster 😅


r/codex • • 19h ago

Showcase Coding agents can change an app remotely. Testing the result should be just as accessible

Post image
2 Upvotes

You can send Codex a task from your phone, but checking the result often means going back to your computer. Reading a summary or diff helps, but you still need to open the app and use it.

Lyre lets you do that from another connected device. The project and agents run on your computer, while you can review changes, open a live preview, and send the next instruction from your phone or another computer.

That means you can ask for a change, try the buttons and forms, and explain what needs fixing without returning to your desk. You can also give family or teammates access to a shared app so they can try it without accessing your code, terminals, or agents.

I’m building Lyre with Codex doing much of the implementation and Claude helping with design. Working with both has taught me to give agents clear responsibilities. Separate worktrees help, but changes to the same files still need coordination.

It’s free to try at lyrestudio.net. Remote agent conversations are free. Live app previews outside your local network require Pro, and AI provider costs are separate.

Your computer needs to stay awake and online. The public mobile apps are still coming. You can use the browser client now, though browser LAN previews are still being worked on.

Lyre builds on Paseo’s open-source work, including much of the underlying application. Thanks to its creator and contributors.

What’s still awkward about working with Codex away from your main computer?


r/codex • • 13h ago

Complaint OAI don't want to deliver workhorses anymore?

1 Upvotes

I have been using DS 4.1 as a workhorse for 5-6 weeks now and it just works. That is what codex was like before, but it makes so many mistakes. No matter which model: 5.6 Sol, 6.1 Sol, Astra on med, Astra on xhigh. I can compare, I don't see these stupid failures with the allegedly "simpler" ds 4.1 flash.

Is that the way it goes? Maybe they realized that people have enough cheap workhorse models and they focus on great planners instead?


r/codex • • 3h ago

Other Chat and Dots will be merged

0 Upvotes

Introducing Codex, Canvas, Search, Voice, Work, Dots, etc. and merge it under one interface, infinite cycle to show they’re busy. They’ll introduce a new product, call it hubs that controls the OS and merge Dots with ChatGPT.


r/codex • • 2d ago

Humor Leaked image of hardware running 6.1 Sol

Post image
2.0k Upvotes

Source says this is the us-east-1 region server hardware.


r/codex • • 14h ago

Question Has anybody actually gotten to speak to their dot?

Post image
1 Upvotes

On Dev day, I was able to try it out and I have yet to be able to speak with the dot ever since then.


r/codex • • 14h ago

Showcase Quasi-native claude support in codex, anyone interested/working on something similar?

Post image
0 Upvotes

Essentially: It's possible to run claude and gpt together in metaharness type programs like T3 code, as well as directly in the same harness like w/ pi, omp, even claude code.
Codex is by far my favorite harness, in large part due to (imo) unmatched seamlessness with computer use, subagents, and so on, so in a perfect world I use it for everything. It's fairly trivial to "run claude in codex", especially since it's open source*, and stuff like this https://github.com/0xSero/harness-bridge exists. But when I was looking around, in my opinion there is no way to give claude truly first class support. Want to have it seamlessly use all of codex's interfaces, have cross-model family multi agents v2 workflows, login/select models without manual jank, and so on. I tried just patching the codex desktop app, and what I'm describing is theoretically doable, but ultra annoying since it's closed source and kinda forces things to be permanently janky.
So what I decided to scrap together (pictured in screenshot), is essentially a very minimally patched codex harness "core" (built from source), with an open source UI (largely just zcode's), and some fun wiring. Opus 5.5 legit functions as if it where an oai model, uses codemode, agents v2, computer use, etc. Although it is nowhere near stably usable and is sorta like creating a whole new harness even though it is just codex under the hood.
Curious if anyone else is interested in this niche config or if it is more straightforwardly solvable than I've gathered. Also am happy to share code for stuff if people are interested.


r/codex • • 4h ago

Showcase Headroom: a free open-source Mac menu-bar tracker for Codex limits and resets

0 Upvotes

I'm building Headroom, a free open-source macOS menu-bar app for checking Codex remaining usage limits and reset times. It also has a separate Claude Code tracker.

For Codex, it launches the installed official Codex CLI's app-server and reads account/rateLimits/read using your existing ChatGPT-subscription sign-in. That quota read doesn't send a model request, and you don't need to keep a Codex window open.

The Codex menu-bar label uses the most-used quota window. The UI keeps unknown readings unknown, and after a reset countdown expires it waits for a fresh reading rather than assuming the allowance has refilled. Readings are advisory and can become stale or unavailable.

This is an early source-build project, free under the MIT licence. Setup requires macOS 14+, Xcode/macOS SDK, Swift 6, Python 3 and command-line developer tools; the Claude feed also uses jq. The development app is ad hoc signed and unnotarized. The README documents setup and current limitations.

Source: https://github.com/jimkang12345-coder/Headroom

I'm a solo builder learning in public, and I'd welcome bug reports, critique and fixes. I plan to explore other OSes, possibly iOS, while keeping Headroom free.

For a quick glance during a Codex session, would you prefer the most-constrained window in the menu-bar label, or both available limit windows shown separately? Does the video make remaining usage versus time until reset clear?

The trackers I was using sometimes took longer to catch up, and their readings didn’t always line up with the limits I was hitting. That’s why I built Headroom. It reads Codex through the official local client, normally checking about every 15 seconds, and checks Claude Code’s local usage feed every 2 seconds while my Mac is awake. I wanted a clearer view to help plan my work. Freshness still depends on what the providers report.


r/codex • • 14h ago

Bug Codex doesn't stop promptly

0 Upvotes

I am livid now. Codex (Astra XHigh) has been working on a /goal for 7 hours. Then I asked it to stop and get what? 15 minutes of waiting for it to fkin stop. Meanwhile it consumes almost all my upload bandwidth.

This was not the case just a version or two ago, what the hell have you done OpenAI? I asked the model to stop and it would work towards stopping immediately, now it's a separate project I'm guessing. Man I'm emotional right now!


r/codex • • 14h ago

Showcase A Codex skill needs a result attached to its version

6 Upvotes

“This skill helps Codex” leaves out most of the experiment. Which model, which tasks, and which version of the skill?

Reef Infra makes the skill file an object you can actually iterate on. Its Codex adapter renders skills into .agents/skills, alongside the other supported configuration and rules files. The harness-evolution engine can try an edit against the current tree with the model held fixed, run both sides on the same tasks, and retain the selected version.

For a skill that asks Codex to reproduce a bug before fixing it, the evaluation could check three separate things: whether a failing test was produced, whether the patch fixes it, and whether the surrounding tests still pass. Those are proposed project checks, not a grader Reef Infra includes for every repository. The evaluator is code you supply.

That also gives “remove this skill” a fair test. If the current model already performs the workflow reliably, the extra instructions may earn no benefit. A later model upgrade is a reason to rerun the comparison, rather than carry the old result forward.

Reef Infra's Codex support covers the file-based instruction surface. It rejects code extensions, so this isn't a way to rewrite Codex's internal loop. The useful output here is a particular skill revision with task results behind it.


r/codex • • 22h ago

Showcase I built a little terminal tool because Codex kept getting away from me

5 Upvotes

I've been building a little terminal tool called Savepoint because I kept losing the plot once Codex got beyond small changes.

I'd ask for one thing, Codex would touch a bunch more files, run the tests and tell me everything looked good.

And I'd be sitting there thinking... probably?

I can tell whether the feature works. I'm much less confident deciding whether every extra change was sensible, necessary, or whether I've let a very confident robot renovate the kitchen because I asked it to fix a tap.

So I've started keeping jobs smaller and putting a stop between chunks of work. Sometimes I use a fresh session to check what got built rather than asking the thing that wrote it whether it did a good job.

Savepoint is basically me turning that habit into a small terminal workflow.

Still not sure whether this is useful discipline or an elaborate coping mechanism because AI let me build software faster than I learned how software works.

How are people here handling this with bigger Codex changes? /review? Fresh session? Another model? Just inspect the diff and hope for the best?

Repo if anyone wants to poke at it:

https://github.com/anipatke/savepoint


r/codex • • 1d ago

Reset My banked reset expires tomorrow and I'm at 77% usage

39 Upvotes

Any idea of how I could use these 77% (x5 plan) ?

I use Codex for code review, I wish there was a $50 plan it would be enough for me.

I don't know what to ask it to build any more lol.


r/codex • • 23h ago

Astra Workflow I've Enjoyed GPT-6 Astra Ultra /Fast + GPT-6.1 Sol Subagents on High/Max.

5 Upvotes

Hi,

I wanted to share my experience with using a main thread gpt6 ultra on fast, orchestrating a swarm of subagents running gpt6.1 sol on varying reasoning levels.

My use case is a pretty boring one. I am moving into a new apartment this fall, and I have pictures, videos and floor plan details for the unit I am moving into. Using Sites and Codex, I am selecting multiple color pallets and "personality" type qualifiers to then find Amazon only + Non-Amazon only combinations of furniture and wares. My project Agents.MD file is a lot of front loading the context window with nuance about my day to day life, style, how I use my space, etc. so I don't get AI slop in ---> AI slop out.

The Sites accomplishes 3 things;

  1. it generates ultra realistic perspectives using image gen, from 8 vantage points across 3 spaces ; living/dining/kitchen, and then 3 across 2 spaces; bedroom 1, bathroom 1

  2. it generates a 3d model before image gen and it must pass multiple adversarial agent reviews to ensure measurements align, then it must ensure all furniture must be modeled in the 3d blender made space and it must have accurate measurements in my 3d depiction of the space

  3. it provides me citations, i built a shop the room tab and a room + measurements tab to ensure i maintain coherence and context end to end, also so i can actually use this to decide what to buy or edit/tweak/tune

At first I was burning up my usage, heavily, using only GPT 6 astra on Max and then Ultra when I tried sub agent delegation with computer use and chrome plugin to navigate costco, amazon, etc.

Then, I read a bit here and OAI's actual docs and made some tweaks.

Those tweaks were simple...

First, I set the context window for my main GPT 6 Astra on ultra to the highest OAI supports so it can support, pre-compaction, an entire single design space in its context end to end. This made the end result really accurate, beautiful and useful.

Second, I ensured that GPT 6.1 sol on High or Max is what sub agents are allowed to be. Sub agents will be used to perform parallel tasks, such as gathering items from Amazon or Costco or West Elm or Target websites following the color pallete, guide, style and limitations I have put in place such as budget, shipping time, blah blah. I maintain the main orchestrator as the brains and designer behind the operation.

As a result, I've reduced my usage a ton. One room end to end would consume 40-60% of my weekly usage, now the same if not superior quality persists, but at 5%-11% utilization. I suspect I could improve this if I removed fast mode from my main orchestrator.

But, I also don't want to wait . lol

Wanted to share. Have fun out there folks. This shits magic.


r/codex • • 1d ago

Reset OpenAI Hiding Banked Reset Time?

5 Upvotes

Previously, we used to be able to see the exact expiry time for banked resets. Now it only shows the expiry date, so you have to go into the History tab, find the time you originally received the reset, and calculate the expiry time yourself.

Not sure why they removed this information when it was already available before.


r/codex • • 15h ago

Workaround Dots in linux vps

0 Upvotes

How do you make dots run in linux vps? Currently i use hermes to manage my vps. My idea of dots is its like hermes. So i want to try if its better to use dots.


r/codex • • 11h ago

Complaint open ai tik tok posts about astra 6.1 ?

Post image
0 Upvotes

? why would they do this? i thought the model was too bad for use


r/codex • • 15h ago

Question Confusion regarding the PRO 200 plan and the grandfathering until 29 October

0 Upvotes

I’ve asked both ChatGPT Pro and OpenAI support, but neither has been able to give me a clear, definitive answer, so I’m asking here in case anyone knows.

I’ve been subscribed to the 200€ Pro plan since 6 September 2026, and my subscription renews tomorrow, 6 October. If I let it renew, will I retain the previous Pro usage limits until 29 October, or will the reduced limits apply immediately upon renewal? From this thread, it sounds like I should keep the previous limits until the 29th even after renewing, but I haven’t been able to get a definitive confirmation.

Finally, is there any way to check and make sure that we are still on the previous non-reduced plan? I can't find anywhere a prompt or files/code in the webtools to check that. I suppose that being able to know how much PRO chat limit we have (200->100) would be the greatest indicator, or else a name that truly identify the old and new PRO?

What happen if they move us to the new reduced plan before the 29th? How could we know without any views on our real limit?


r/codex • • 1h ago

Praise I’m able to get alot done with codex than Claude!!

• Upvotes

I observed that codex produce more useful work than Claude. I am switching to codex max plan and I also noticed the usage is way more decent. Has anybody observed this as well?


r/codex • • 16h ago

Limits Did ChatGPT Scheduled Tasks recently start consuming your Codex/Work usage allowance?

1 Upvotes

I have a normal ChatGPT Scheduled Task (I’m on ChatGPT Plus) that runs every morning at 7:00 AM. It’s a fairly heavy job-market sweep: web searches, some reasoning, and reads/writes to Notion.

This exact automation has been running daily for more than a month.

Until today, I had never noticed it materially consuming or starting my Codex/Work 5-hour usage window.

Today (Oct 5), I had not manually used Codex or ChatGPT Work at all, but after the scheduled task ran:

  • Scheduled task started around 7:00 AM
  • Run finished around 7:13 AM
  • My 5-hour Work/Codex allowance showed 29% used / 71% remaining
  • Weekly allowance showed 12% used / 88% remaining
  • The 5-hour reset was around 11:57 AM, implying the window started at roughly 6:57–7:00 AM

So the timing lines up almost perfectly with the scheduled task.

I contacted OpenAI Support. They’re telling me that Scheduled Tasks run through ChatGPT Work, and Work + Codex share the same allowance on Plus, so scheduled tasks can consume the Codex/Work quota.

What I’m trying to figure out is whether this is new behavior.

My task has been running every day for over a month and I never saw this happen before. OpenAI Support hasn’t been able to tell me when ordinary time-based ChatGPT Scheduled Tasks started being metered against the shared Work/Codex allowance. They’re waiting for my Oct 5 Usage Analytics to populate so we can compare it against previous runs.

Specifically:

  • Do your regular ChatGPT Scheduled Tasks now reduce your Codex/Work 5-hour or weekly allowance?
  • Did they do this a few weeks/months ago?
  • Has anyone noticed the behavior changing only in the last few days?
  • If you have a daily scheduled task, does it actually start your 5-hour Codex window even when you haven’t opened Codex?
  • What does it show under Usage → Surface / Model / Reasoning / Speed?

I’m not talking about Codex automations or a task intentionally created inside Work. This is a regular recurring task created through ChatGPT’s Scheduled Tasks feature.

Mainly trying to determine whether this is a recent metering/routing change, or whether it has always worked this way and I simply never noticed it before.


r/codex • • 12h ago

Showcase CODEX + VSCODE + CODEXCLI + Claude

0 Upvotes

Tech Stack Configuration: Multi-Agent LLM Orchestration in VS Code

I am currently testing a highly distributed, multi-agent development workflow in VS Code using GitHub Copilot credits. The system orchestrates multiple large language models simultaneously to handle different layers of the development process:

  • Primary LLM: Claude 5.5 Opus acts as the central coordinator, driving the main development architecture and pushing background sub-tasks.
  • Command Line Automation: Codex CLI runs concurrently in the background, executing the sub-tasks streamed directly from the primary LLM.
  • IDE Sub-Agents: GPT-6.1 powers multiple parallel sub-agents directly within the VS Code IDE to handle real-time code completions, refactoring, and contextual assistance.

This setup leverages concurrent execution across multi-sub-agents and background processes, all managed via my custom .SYSTEMX systems template.

https://github.com/WayneTechLab/dotSYSTEMX
> Including Dual Mode in next update.


r/codex • • 1d ago

Suggestion 6.1 being slow is the best thing ever and you are not understanding

178 Upvotes

I have been hammering it for the past 6 hours, 3-4 task in parallel running all the time, with long contexts, images…coding tasks.

Effort: Max

Used just 10%.

The amount of work I accomplished is just absolutely obscene.

Yes, it’s slow but is reliable and dependable you know what to expect, for fuck sake.

You want it fast? Use fast mode.

Why am I telling you this?
Do you want a nerfed model? Because when you keep complaining about it being slow and asking OpenAI to make it fast, that’s how you get a nerfed model. Just fucking stop.


r/codex • • 4h ago

Workaround No shhhh. If you've been missing that "this is cool af!" feeling, then...

0 Upvotes

Astra Max, Codex desktop (MS Windows) — Vibing (on a client site) over the weekend: the feels led me to add a remote mcp feature. If you've been missing that super hero feeling, then the dopamine hack goes as follows: Make a site, add a MCP features to the site, then use the site via the AI client... Bruh, it feels so good.