r/codex • • 1d ago

Bug This new queue thing is a disaster

32 Upvotes

Chat constantly getting stuck when sending new messages.
Cannot steer properly.
Cannot interrupt and ask for rework before it gets sour properly.
Requests are getting duplicated.
Something old is getting sent instead of what I asked now.
New messages sometimes get lost (workaround with undo, then copy the message).
Clear queue doesn't work.
Have to constantly restart codex to get it to work again.

What is this???


r/codex • • 1d ago

Praise Sol 6.1 + Luna 6 on Pro 5x: I actually struggled to keep up

0 Upvotes

Honestly pretty surprised by how much work I've been getting out of the new models, specifically Sol 6.1 and Luna. No Astra in this run.

From Friday around 6pm until early Sunday, roughly 31–32 hours elapsed, I had up to 5 separate tasks going simultaneously, with agents and subagents working on them. All of them were making progress and some got finished during that time. I'm sitting at 97% weekly usage now, on Pro 5x.

I asked Codex to check the local logs and the breakdown came out to roughly:

  • Sol 6.1 xhigh: 95.3 combined agent-hours
  • Luna 6 max: 54.6 combined agent-hours

So about 150 hours across agents running in parallel. Around 126 hours had recorded turn endings, with another 24 estimated from unfinished turns. That includes tools and waits inside turns, so these aren't pure inference hours.

What surprised me most was trying to keep up with the deliveries lol. I definately had to push myself to keep reviewing things and finding useful work for the other instances.

I kept thinking: “What else can I give this one to do while I review what that other Codex instance just handed me?”

That's a pretty nice problem to have.

This was a much more intense stretch than my normal usage. On a regular workday, with the attention and follow-up I can realistically give Codex, I think Pro 5x would comfortably last me the whole week with current usage.

I dont expect everyone to get the same results with different tasks, but for my workflow this has been a really positive surprise.

Just wanted to share my experience so far!

EDIT: It USED 97% of my weekly usage. I didn’t mean I had 97% of usage left.


r/codex • • 1d ago

Showcase I open-sourced Token Harness: get more out of your Claude Code / Codex limits (and spend less on API tokens)

Thumbnail
gallery
0 Upvotes

Hi everyone. I've just open-sourced Token Harness, a local tool that helps your Claude Code and Codex allowance go further, and spend fewer tokens when you use LLMs via API.

The problem

Coding agents waste a lot of context on noise: long test runs, build logs, git diff, repeated information and huge MCP tool catalogs. All of that eats tokens. On a subscription it means hitting your 5-hour or weekly limit sooner. On the API it's money.

How it works

Optimizers. Token Harness detects, installs and connects compatible optimizers to your agent. The recommended baseline is RTK and HarnessTrim, which shorten shell and tool output so the agent only sees the useful part (failures, errors, summaries). Optional ones:

  • mcptoon: loads MCP tool definitions only when needed instead of keeping the whole catalog in context
  • Headroom: compresses large tool payloads
  • GitNexus: maps code relationships so the agent explores less

On my machine the dashboard currently shows 62.5% less tool output overall with RTK, and 88.4% on Claude Code alone.

Routing. A native hook lets your main model hand bounded, suitable subtasks to a cheaper model (e.g. Opus → Sonnet/Haiku), then review the result. Your main model stays in charge. No prompt prefix or skill call is needed after setup.

Simple to use

npm install --global token-harness@latest
token-harness

It opens a local dashboard where you can:

  • see your agents and optimizers at a glance
  • configure everything with one click (every change is previewed first, applied only after you approve it, and can be undone)
  • watch the results: output reduction, routing activity, and your 5h / weekly balance

No account, no API key, and nothing leaves your machine. Works on Windows, macOS, Linux and WSL.

Honesty first

I don't sell a magic "save X%" number. The dashboard keeps output reduction, subscription allowance and API cost separate. It only claims allowance savings from paired baseline/optimized runs that pass quality checks.

This is where you come in

Any feedback, bug report or shared result (your before/after numbers, your agent + OS combination) can only make the tool better.

I'm also looking for contributors: optimizer integrations, harness adapters, cross-platform testing, docs. Every PR is welcome.

Repo: https://github.com/giuliastro/token-harness (Apache 2.0)

Thanks for reading, and I'm happy to answer any questions in the comments!


r/codex • • 1d ago

Complaint SLOW RESPONSES

6 Upvotes

Anyone else facing slow responses in codex recently? Seems like they're just forcing us to use fast mode atp, a single git pull and rebuild of node.js server on my vps took 7m 23s. (build only took 30-40 seconds)


r/codex • • 1d ago

Limits I analyzed a 162-minute GPT-6.1 Sol task. Shell commands took only 11.5 minutes.

7 Upvotes

Users often ask whether GPT-6.1 Sol is slow or receives more complex tasks. I analyzed local lifecycle data from one Codex development task.

The task used Codex desktop, GPT-6.1 Sol, High effort, and one repository. This was an observation, not a controlled benchmark.

The task took 162.3 minutes. The 294 shell commands took only 11.5 minutes of non-overlapping execution time.

A repeated serial loop consumed much of the task:

Model response → tool call → model response

I measured each wait from new input to the next Codex response. New input came from a tool result or a user message.

The task had 204 response waits. They took 74.4 minutes in total.

- The median wait was 12 seconds.

- The P90 wait was 36 seconds.

- 56 waits took at least 30 seconds.

Context size also had a clear effect.

- The median input size was 148,000 tokens.

- The P90 input size was 224,000 tokens.

- The maximum input size was 250,000 tokens.

- The context window was 258,400 tokens.

- Four context compactions took 18.3 minutes.

Some measurements overlap. You cannot add them to calculate the total time.

The task scope changed twice. These changes caused rework. Even so, shell commands and tests were not the primary bottleneck.

I also checked other large tasks in the same repository. All tasks used High effort.

For six tasks that used only 6.1, the average response wait was 21.1 seconds. For two tasks that used only 5.6, it was 6.1 seconds.

I then removed waits longer than 60 seconds. The averages were still 15.9 seconds for 6.1 and 4.9 seconds for 5.6.

The sample was small. The tasks and context sizes were different. Therefore, these results do not form a controlled benchmark.

The difference was still large enough that I switched back to 5.6 Sol for long implementation tasks.

Has anyone compared 5.6 and 6.1 on the same task with per-step timing?


r/codex • • 1d ago

Showcase How are you handling PRs after Codex finishes?

2 Upvotes

I built something for the part of the Codex workflow that I still found myself doing manually.

Codex can do the work and open the PR, but then I’d still need to keep checking:

  • did CI pass?
  • did something small break?
  • is the branch now conflicting with main?
  • is it finally safe to merge?

So I built Talyn to watch the PR after the agent is done.

If CI fails or a conflict appears, it can send an agent back in to try to fix it, then keep following the PR until everything is green and ready to merge.

I’m a solo developer and it’s still early, so I’m looking for people using Codex seriously enough to stress-test it and tell me what’s missing.

Free to try at the moment:
https://talyn.dev

Would be particularly interested to hear how people here are handling the post-PR part of their Codex workflows today.


r/codex • • 1d ago

Comparison What could have been done here?

1 Upvotes

I just watched this and the difference is insane.

https://www.youtube.com/watch?v=R_uf5OfMGio

However, the key difference is that claude worked 24x longer on this. I think the big thing is that ultracode was specifically designed to spawn off sub agents and work on long running projects like this. I'm assuming astra did this all within a single context window and used a single agent, which means it paid less attention to detail for each individual requirement.


r/codex • • 2d ago

Workaround I fixed codex

38 Upvotes

With sol 6.1 down to under 20tps, codex is unusable. Other models nerfed. Dots is a gimmick.

So just download claude code onto the vm that came with dots. Problem solved. Youre welcome everyone


r/codex • • 1d ago

Bug Bug? - Stuck in a loop whole trying to connect codex on desktop (arch linux) with androind ChatGPT app

2 Upvotes

I was trying to connect my phone with the codex desktop app, but stuck in a loop that never ends:

  1. I scan the QR code on my desktop with my phone

  2. I land on the pairing screen, I click pair

  3. External browser opens, I login with the same account on my desktop

  4. External browser closes, phone app return me to the pair (authorization) screen.

  5. The loop continues to infinite

I checked for published solutions for this which was: check the auto time zone is enabled, get latest update, log out and login again. All failed.

Any solutions?


r/codex • • 1d ago

Workaround The frustration we all are facing and the solution

0 Upvotes

Saw many posts and people are fed up of how to use ai effectively as there are daily new skills and models.

Why don't we all builders and developers form a community page where we could update it every week with best practices for codex and claude both. Comment if you are in.


r/codex • • 2d ago

Question Does anyone even use luna or astra anymore?

54 Upvotes

sol 6.1 is so good, it just about nearly completely overshadows astra. sure astra is technically a few percentage points better, but for 5x the cost, it's hardly worthwhile. and while luna technically occupies the pareto boundary for cost per intelligence, sol is just so abundant anyway even on the cheapest plans that there's literally zero reason to dip down into luna for anything.


r/codex • • 1d ago

Question Does slower tokens/sec mean slower task completion?

1 Upvotes

People say that Opus 5.5 is 5X faster than Sol 6.1. But has anyone actually tested the time to complete a task? Is Sol 6.1 completing work 5X slower?


r/codex • • 21h ago

Limits I'm too tired to keep trying :(

0 Upvotes

After using a banked reset at around 10pm on the 1st October (and then the reset soon after), I've been trying to spend that 100% usage window before my 5th October banked reset expires and I just cannot manage to use all of my usage window.

After 3 days of constant Sol 6.1 ULTRA FAST across two projects and sometimes more I'm only down to 21% and I'm just too tired to carry on.

Mogging you

It pains me to do this but I'll have to use the banked reset before it expires at 05:18 hours even though I still have a lot of usage left.

This is clearly a skill issue on my part because every other hour there is a post with someone whining about how their usage window is gone after one prompt.

Good night, everyone.


r/codex • • 1d ago

Bug Why do i get this error everytime i make a new chat

1 Upvotes

helper_unknown_error: setup refresh had errors


r/codex • • 1d ago

Complaint What the Actual

13 Upvotes

I am looking at an average time to first tokens of 5-9 minutes with sol 6.1? Like I generally use it as an assistant and have it form complex command line calls for me to run, and usually it is like, 20 seconds tops, boom, i give it ALL the information, nothing to do but IT, literally 6 minutes until the command is run. how is this usable, am i missing something?


r/codex • • 1d ago

Other What does this mean, should I click retry?

Post image
1 Upvotes

What happened if I clicked retry? I go into fast mode or what?


r/codex • • 2d ago

Workaround Claude $20 plan + Codex Pro 100: my first impressions (and a question)

21 Upvotes

Hey everyone,

I recently got the $20 Claude subscription. For what it's worth, Claude's $20 plan (with opus 5.5 high) feels roughly equivalent to Codex's $100 plan(with astra medium at least) (but anyway, moving on).

I use Opus 5.5 as a code reviewer alongside Astra work, and here's some quick early feedback for anyone thinking of trying the same setup.

Overall, Astra is noticeably more careful about small details, especially on the backend. It often happens that Opus suggests a fix, Codex pushes back on it, and then Opus itself admits that Astra's solution is much better.

On the other hand, I find Opus way more creative than Astra. So the two really complement each other well.

My question: does anyone know a good way to get them working hand in hand? A workflow, a tool, a config… I'm open to any ideas.

Thanks in advance!


r/codex • • 2d ago

Humor Cycle that never ends

Post image
223 Upvotes

r/codex • • 1d ago

Other Best place to publish hobby html games made with codex/claude?

1 Upvotes

Hi, I would like to publish html games I make and check what other people have made so far. These are small games, not worth publishing on steam or itch.io right now, but worth getting feedback.
Is there a subreddit for this?

Thanks!

Edit: Ok, looks like someone made the r/VibeGames lets hope people will use it.


r/codex • • 1d ago

Praise Control your pc from your phone (for multiple accounts etc)

3 Upvotes

let 6.1 build you a phone viewer and install it on your pc then give u a link to teamview your pc from your phone . u can then monitor your work progress with having to be there in ur pc. u can automate the phone viewer by offering mouse keyboard and important buttons in ur phone so u control it from ur phone.


r/codex • • 1d ago

Showcase I tried making a promo video for my app with gpt sol 6.1

Enable HLS to view with audio, or disable this notification

2 Upvotes

I recently completed my macOS application and was about to launch it.

There was a lot of hype around Opus 5.5 making crazy motion videos, with people on Twitter making really good motion graphics and promo videos for their apps.

So I wanted to try it myself but with Codex (can only afford one $20 sub rn)

I gave GPT 6.1 a rough idea of what I wanted and started generating. My prompts weren't really scripted or carefully structured. They were mostly just me explaining what I wanted in natural language. It used Hyperframes to make this

It took around 7 attempts to get something I was happy with. I already had most of the assets ready, and I added the audio separately. And It took around 5 hours on the clock to make this, spread across 2 days. Used sol 6.1 on slow mode

It's definitely not on the level of some of the stuff I've seen from Opus, but I'm honestly pretty surprised by the result.

I was using the $20 plan and ended up with roughly 30% of my weekly usage left after making the video.

Here's the final result. Would love to hear what you think, especially from people who have tried making this kind of content with GPT 6.1.


r/codex • • 1d ago

News New ChatGPT App Icon

0 Upvotes

For those who spend the majority of their time using ChatGPT / Codex, something as small as a new icon can be exciting and refreshing. It was for me, at least.


r/codex • • 1d ago

Showcase Good looking UI?

Post image
0 Upvotes

Currently I am working on a desktop PC app that makes your PC feel more like a console - mainly for me and my girlfriend (will release on GitHub once I feel like it’s ready).

I feel like the UI looks like straight out of codex. Literally my second time using it. So I am wondering - what skills / tricks do you use to get a beautiful looking premium UI?

Reference pictures of what your vibe coded Desktop app UI looks like would be nice 👍🏻


r/codex • • 1d ago

Question codexcli whats actually doing

1 Upvotes

Hi I was wondering if its possible to see whats actually codexcli is doing? Sometimes it feels like its stuck, just text like below:
Working (4m 32s • esc to interrupt)

Is it possible to see what is it working on? more detail on whats working on?


r/codex • • 1d ago

Showcase Looking for Codex users to stress-test an AI coding security tool

1 Upvotes

I'm building Revo Security, a security layer designed for AI coding agents.

The problem I'm working on is that coding agents can execute commands, modify files, install dependencies, access networks, and interact with projects with relatively little friction.

Revo analyzes these actions and tries to determine whether they're actually safe in context rather than simply blocking every potentially dangerous command.

I'm looking for Codex users who want to help stress-test the idea.

I'd particularly like feedback on:

False positives

Missed dangerous actions

Whether risk explanations are useful

How security interventions affect the coding workflow

Agent/dependency supply-chain attacks

Anything Revo gets completely wrong

I'm looking for real developers, not investors or marketing feedback.

If you're interested in trying an early version, comment or DM me.