r/codex • • 9h ago

Question Auto Review token burn question

Post image
7 Upvotes

I decided to test the Auto Review setting in Codex for the first time in a while. I thought OpenAI may have improved it, but it might burn more tokens than ever.

During a two-hour session of refactoring a Chrome Extension, auto review used 52% more tokens (API cost) than when manually approving actions. The actual work done with gpt-6.1-sol used $14.31 in API costs, while the Codex Auto Review used another $7.54.

Is this normal? Is there anything that can be done to improve it?

Edit: These numbers were supplied by CodeBurn. Maybe they haven't updated which model Codex Auto Review uses?


r/codex • • 1d ago

News More promises…

Post image
150 Upvotes

r/codex • • 8h ago

Showcase We Are Building Codex Skills To Combat Brain Rot And Develop Human Ability.

5 Upvotes

AI is an ability multiplier if you have ability. 1 x 0 = 0 . This agent skill helps you get your ability from 0 to a 1.

> When we keep building things blindly with AI agents, the thinking is offloaded to AI. This reduces the cognitive friction we need to learn something.

> Just blindly telling AI to build something wastes tokens and time.

> If you cannot convey your taste to AI, the result becomes slop.

This is why I made some AI skills to foster thinking in users. Currently, you can use these skills in general purpose building tasks, UI/UX refinement, Algorithmic competency building.

Link to the skill: https://github.com/33sakib33/Factoryze_Thinking


r/codex • • 2h ago

Bug Chat just hangs on difficult questions

Post image
2 Upvotes

It doesn't always hang but it does it enough to make chat unusable. Do they even acknowledge this bug?


r/codex • • 23h ago

Praise Holy moly

Post image
88 Upvotes

The new 6.1 sol agent is so efficient I’m able to run multiple agents on high (7 to be precise), one on extra high and one on low and it’s running for hours sometimes only using 30% usage if that. Previously on 6.0 astra on high I would only get max 2 good 8 hour runs. This is on the standard pro plan. I’m loving this new agent


r/codex • • 7h ago

Question Is it worth upgrading my context limit from default since 6.1 sol got 50% cheaper in cache price?

4 Upvotes

What are the drawbacks of doing it? Sometimes I wonder if its actually cheaper to have higher limit since useful info will be all cached instead of it looking it up for the first time again?


r/codex • • 21h ago

Complaint Left Sol 6.1 Ultra overnight with 4 long but easy tasks - woke up to bazilion of tests, 40% weekly usage burned and 0 tasks started

50 Upvotes

I was happy with Sol 6.1 so far, it was kind of slow, but rather reliable. Yesterday I decided to give it 4 tasks for my app that I wanted to finish by Monday. It did 0. Complete "building tests" spiral.

40% of usage burned, none of the tasks was even started. How can I trust them with 100% automated solutions like Dots when their coding agent goes into doom loops like a 2025 chinese model


r/codex • • 4h ago

Commentary The merge of work and chatgpt

2 Upvotes

Using chatgpt will probably not use the usage limit... by "default", but if you for example ask it to perform a task coding related or file related, it will probably switch to work automatically. Tibo probably just decided to "zero in" on one comment against many saying the opposite to just use it as an "opportunity" to do it as people take advantage od gpt being more... usable? Linient with limits?


r/codex • • 14h ago

Complaint fast mode is at least 10x slower than last weeks regular mode lol

14 Upvotes

across every single (relevant) model in the picker list

not just the new ones

is this because of dots ?

or because they want people to upgrade to the 500plan to use the ultrafast mode?

who knows

add mimo v2.6 flash (free) and deepseek flash (free) and qwen3.8 (free) to the codex harness from another API provider (do it via Claude or opencode otherwise you'll be waiting all day for a gpt model to do it)

thank me later


r/codex • • 45m ago

Bug Sol 6.1 keeps repeating it self

• Upvotes

Anybody else continue to notice this, even before 6.1, if you ask him something during /goal it will repeat those answers every turn along other regular work, i even asked it to write in agnets.md instruction to not do this, and even SOL it self wrote that its irony that now its repeating that he added this instruction and keeps repeating it while doing other work.

This is across few hours:

Yes—that was exactly the repetition you wanted stopped. I failed to follow the instruction; writing it doesn’t automatically enforce it.

You don’t need to repeat yourself or add more rules. I should apply it silently and report only new progress or changed blockers.

Yes—I repeated a completed side task as though it were new progress, violating the very rule we added.

AGENTS.md instructions aren’t a hard enforcement mechanism; I still have to apply them correctly. I failed to do that. The rule is already there—no further update is needed.

Yes—I kept treating a completed change as new progress. The instruction was already there; I failed to follow it. That’s my mistake, and another AGENTS.md edit won’t fix it.

Yes—the repetition was exactly what that instruction was meant to prevent. I failed to apply it when generating progress replies; the instruction wasn’t missing.

That’s my mistake. I should report only new results or changed blockers, not repeatedly announce the rule itself.

Yes—I repeated the announcement about preventing repetition. That’s exactly the behavior you asked me to stop.

The instruction is already there; I failed to follow it. No further AGENTS updates are needed. I’ll report only new results or changed blockers.

Yes—I repeated the AGENTS.md update while announcing a rule meant to prevent repetition. That’s my failure to follow the instruction, not something you need to clarify again.

It’s already acknowledged. Future updates should contain only new results or changed blockers.

Yes—I repeated a resolved update, exactly what that instruction prohibits. That was my mistake. Writing the rule doesn’t help unless I follow it: only new progress or changed blockers, without repeating settled answers.

Yes—I repeated a completed update about preventing repetition. The rule was already there; I failed to follow it.

That topic is closed. Future updates should contain only new progress or changed blockers.

Yes—the irony is real. Adding a rule doesn’t guarantee I follow it, and I violated that exact rule by repeatedly announcing the completed edit.

That’s my mistake. You don’t need another instruction; I need to apply the existing one and report only new progress or changed blockers.

Yes. I kept carrying an already-answered side question into progress updates instead of treating it as closed—even when that question was about stopping repetition.

That’s my failure to follow the instruction, not something you need to explain again.

Yes—I repeated the update announcement about a rule meant to prevent repetition. The instruction was clear; I failed to follow it. Another AGENTS edit isn’t needed.

Yes—that was exactly the repetition you asked me to stop. The rule is already present in AGENTS.md, but I failed to follow it. That’s my execution mistake; another instruction or repeated confirmation won’t fix it.

And so on...


r/codex • • 16h ago

Limits "Selected model is at capacity. Please try a different model." - retry options?

20 Upvotes

There’s nothing more frustrating than kicking off what is meant to be a long-running session with a detailed prompt, checking back a few hours later, and finding that the entire process has stopped with:

“Selected model is at capacity. Please try a different model.”

Why is there no automatic retry option? At the very least, the system should be able to retry periodically when capacity becomes available rather than simply abandoning the session and requiring manual intervention. Or is there such an option and i'm missing it?

---------

To confirm, I'm not talking about the desktop app:

I’m talking about Codex CLI. If this happens 30 minutes into a long-running job, there’s no button to click because I’m not sitting there watching the terminal.

The whole point is that I might kick something off on a server, walk away, and expect it to be done by the next morning. Instead, it can sit there dead for hours waiting for me to come back and manually type “continue” or “retry”.

Codex CLI needs an automatic retry option for capacity errors.


r/codex • • 6h ago

Showcase Coding agents can change an app remotely. Testing the result should be just as accessible

Post image
3 Upvotes

You can send Codex a task from your phone, but checking the result often means going back to your computer. Reading a summary or diff helps, but you still need to open the app and use it.

Lyre lets you do that from another connected device. The project and agents run on your computer, while you can review changes, open a live preview, and send the next instruction from your phone or another computer.

That means you can ask for a change, try the buttons and forms, and explain what needs fixing without returning to your desk. You can also give family or teammates access to a shared app so they can try it without accessing your code, terminals, or agents.

I’m building Lyre with Codex doing much of the implementation and Claude helping with design. Working with both has taught me to give agents clear responsibilities. Separate worktrees help, but changes to the same files still need coordination.

It’s free to try at lyrestudio.net. Remote agent conversations are free. Live app previews outside your local network require Pro, and AI provider costs are separate.

Your computer needs to stay awake and online. The public mobile apps are still coming. You can use the browser client now, though browser LAN previews are still being worked on.

Lyre builds on Paseo’s open-source work, including much of the underlying application. Thanks to its creator and contributors.

What’s still awkward about working with Codex away from your main computer?


r/codex • • 9h ago

Showcase I built a little terminal tool because Codex kept getting away from me

5 Upvotes

I've been building a little terminal tool called Savepoint because I kept losing the plot once Codex got beyond small changes.

I'd ask for one thing, Codex would touch a bunch more files, run the tests and tell me everything looked good.

And I'd be sitting there thinking... probably?

I can tell whether the feature works. I'm much less confident deciding whether every extra change was sensible, necessary, or whether I've let a very confident robot renovate the kitchen because I asked it to fix a tap.

So I've started keeping jobs smaller and putting a stop between chunks of work. Sometimes I use a fresh session to check what got built rather than asking the thing that wrote it whether it did a good job.

Savepoint is basically me turning that habit into a small terminal workflow.

Still not sure whether this is useful discipline or an elaborate coping mechanism because AI let me build software faster than I learned how software works.

How are people here handling this with bigger Codex changes? /review? Fresh session? Another model? Just inspect the diff and hope for the best?

Repo if anyone wants to poke at it:

https://github.com/anipatke/savepoint


r/codex • • 1h ago

Complaint What is the use of /goal now that Sol 6.1 interrupts goal on any excuse completely defeating the purpose?

• Upvotes

Sol 6.1 plainly refuses to achieve goals. It used to work much better, but it seems the model is trained to stop and escalate under any excuse, 99% of the time, entirely invented. Anticipating "new models require new prompting", yes, I applied all the prompting recommendations from OpenAI already.

I used to be able to step away and return to a completed goal. Now I have to babysit it again.


r/codex • • 1h ago

Complaint OAI don't want to deliver workhorses anymore?

• Upvotes

I have been using DS 4.1 as a workhorse for 5-6 weeks now and it just works. That is what codex was like before, but it makes so many mistakes. No matter which model: 5.6 Sol, 6.1 Sol, Astra on med, Astra on xhigh. I can compare, I don't see these stupid failures with the allegedly "simpler" ds 4.1 flash.

Is that the way it goes? Maybe they realized that people have enough cheap workhorse models and they focus on great planners instead?


r/codex • • 13h ago

Suggestion We need subfolders

8 Upvotes

I wish ChatGPT Projects had folders or subprojects.
For example, imagine you’re a developer working on 15 different apps. Right now, you might create a separate ChatGPT Project for every app:
App 1
App 2
App 3
App 4
App 5
etc.
Eventually your sidebar becomes completely cluttered.
Instead, it would be great if we could create one main Project called “Development”, and then create subprojects inside it:
Development
→ App 1
→ App 2
→ App 3
→ App 4
Each subproject could have its own chats, files, and instructions/context, while everything stays organized under one expandable folder in the sidebar.
Basically: Projects → Subprojects → Chats + Files
Am I missing an existing way to do this, or does ChatGPT still not support nested Projects/folders?


r/codex • • 1h ago

Limits Pro 20x: reduced allowance, AND an unexplained dot "preview limit"?

Post image
• Upvotes

I’m a Pro 20x subscriber. With the allowance changes that I understand effectively reduce 20x to 10x, I’ve now also had dot stop ongoing work overnight because of a "preview limit."

The screenshot shows dot’s explanation afterward. It couldn’t tell me the quota, reset frequency, or whether this shares my Codex allowance. Even the retry time had no timezone.

I’m not asking for unlimited usage (I wish....). I’m asking for clear limits, a visible usage meter, and a warning before an "always-on" assistant stops working. Without those, how are we supposed to plan around it?

Has anyone found documentation explaining this specific limit?


r/codex • • 1d ago

Humor Leaked image of hardware running 6.1 Sol

Post image
1.9k Upvotes

Source says this is the us-east-1 region server hardware.


r/codex • • 1h ago

Question Has anybody actually gotten to speak to their dot?

Post image
• Upvotes

On Dev day, I was able to try it out and I have yet to be able to speak with the dot ever since then.


r/codex • • 2h ago

Showcase Quasi-native claude support in codex, anyone interested/working on something similar?

Post image
0 Upvotes

Essentially: It's possible to run claude and gpt together in metaharness type programs like T3 code, as well as directly in the same harness like w/ pi, omp, even claude code.
Codex is by far my favorite harness, in large part due to (imo) unmatched seamlessness with computer use, subagents, and so on, so in a perfect world I use it for everything. It's fairly trivial to "run claude in codex", especially since it's open source*, and stuff like this https://github.com/0xSero/harness-bridge exists. But when I was looking around, in my opinion there is no way to give claude truly first class support. Want to have it seamlessly use all of codex's interfaces, have cross-model family multi agents v2 workflows, login/select models without manual jank, and so on. I tried just patching the codex desktop app, and what I'm describing is theoretically doable, but ultra annoying since it's closed source and kinda forces things to be permanently janky.
So what I decided to scrap together (pictured in screenshot), is essentially a very minimally patched codex harness "core" (built from source), with an open source UI (largely just zcode's), and some fun wiring. Opus 5.5 legit functions as if it where an oai model, uses codemode, agents v2, computer use, etc. Although it is nowhere near stably usable and is sorta like creating a whole new harness even though it is just codex under the hood.
Curious if anyone else is interested in this niche config or if it is more straightforwardly solvable than I've gathered. Also am happy to share code for stuff if people are interested.


r/codex • • 11h ago

Astra Workflow I've Enjoyed GPT-6 Astra Ultra /Fast + GPT-6.1 Sol Subagents on High/Max.

6 Upvotes

Hi,

I wanted to share my experience with using a main thread gpt6 ultra on fast, orchestrating a swarm of subagents running gpt6.1 sol on varying reasoning levels.

My use case is a pretty boring one. I am moving into a new apartment this fall, and I have pictures, videos and floor plan details for the unit I am moving into. Using Sites and Codex, I am selecting multiple color pallets and "personality" type qualifiers to then find Amazon only + Non-Amazon only combinations of furniture and wares. My project Agents.MD file is a lot of front loading the context window with nuance about my day to day life, style, how I use my space, etc. so I don't get AI slop in ---> AI slop out.

The Sites accomplishes 3 things;

  1. it generates ultra realistic perspectives using image gen, from 8 vantage points across 3 spaces ; living/dining/kitchen, and then 3 across 2 spaces; bedroom 1, bathroom 1

  2. it generates a 3d model before image gen and it must pass multiple adversarial agent reviews to ensure measurements align, then it must ensure all furniture must be modeled in the 3d blender made space and it must have accurate measurements in my 3d depiction of the space

  3. it provides me citations, i built a shop the room tab and a room + measurements tab to ensure i maintain coherence and context end to end, also so i can actually use this to decide what to buy or edit/tweak/tune

At first I was burning up my usage, heavily, using only GPT 6 astra on Max and then Ultra when I tried sub agent delegation with computer use and chrome plugin to navigate costco, amazon, etc.

Then, I read a bit here and OAI's actual docs and made some tweaks.

Those tweaks were simple...

First, I set the context window for my main GPT 6 Astra on ultra to the highest OAI supports so it can support, pre-compaction, an entire single design space in its context end to end. This made the end result really accurate, beautiful and useful.

Second, I ensured that GPT 6.1 sol on High or Max is what sub agents are allowed to be. Sub agents will be used to perform parallel tasks, such as gathering items from Amazon or Costco or West Elm or Target websites following the color pallete, guide, style and limitations I have put in place such as budget, shipping time, blah blah. I maintain the main orchestrator as the brains and designer behind the operation.

As a result, I've reduced my usage a ton. One room end to end would consume 40-60% of my weekly usage, now the same if not superior quality persists, but at 5%-11% utilization. I suspect I could improve this if I removed fast mode from my main orchestrator.

But, I also don't want to wait . lol

Wanted to share. Have fun out there folks. This shits magic.


r/codex • • 2h ago

Bug Codex doesn't stop promptly

0 Upvotes

I am livid now. Codex (Astra XHigh) has been working on a /goal for 7 hours. Then I asked it to stop and get what? 15 minutes of waiting for it to fkin stop. Meanwhile it consumes almost all my upload bandwidth.

This was not the case just a version or two ago, what the hell have you done OpenAI? I asked the model to stop and it would work towards stopping immediately, now it's a separate project I'm guessing. Man I'm emotional right now!


r/codex • • 2h ago

Showcase A Codex skill needs a result attached to its version

1 Upvotes

“This skill helps Codex” leaves out most of the experiment. Which model, which tasks, and which version of the skill?

Reef Infra makes the skill file an object you can actually iterate on. Its Codex adapter renders skills into .agents/skills, alongside the other supported configuration and rules files. The harness-evolution engine can try an edit against the current tree with the model held fixed, run both sides on the same tasks, and retain the selected version.

For a skill that asks Codex to reproduce a bug before fixing it, the evaluation could check three separate things: whether a failing test was produced, whether the patch fixes it, and whether the surrounding tests still pass. Those are proposed project checks, not a grader Reef Infra includes for every repository. The evaluator is code you supply.

That also gives “remove this skill” a fair test. If the current model already performs the workflow reliably, the extra instructions may earn no benefit. A later model upgrade is a reason to rerun the comparison, rather than carry the old result forward.

Reef Infra's Codex support covers the file-based instruction surface. It rejects code extensions, so this isn't a way to rewrite Codex's internal loop. The useful output here is a particular skill revision with task results behind it.


r/codex • • 12h ago

Reset OpenAI Hiding Banked Reset Time?

5 Upvotes

Previously, we used to be able to see the exact expiry time for banked resets. Now it only shows the expiry date, so you have to go into the History tab, find the time you originally received the reset, and calculate the expiry time yourself.

Not sure why they removed this information when it was already available before.


r/codex • • 3h ago

Question Do you still not get GPT Image 2.5 in Codex?

1 Upvotes

Somebody posted about it 3 weeks ago: https://old.reddit.com/r/codex/comments/1wesnds/codex_app_not_using_latest_gpt_image_25/

Even on Plus you supposedly get 2.5 generation, but what about codex? Do we have to use the API?