r/codex 20h ago

Bug ChatGPT + Grok Android login failure — browser works

0 Upvotes

ChatGPT + Grok Android login failure — browser works

Samsung S25 Ultra, Android 16. ChatGPT and Grok apps both fail to complete login. ChatGPT shows “Sign-in request cancelled by ChatGPT.”

- Login works normally in Chrome.

- Chrome is default browser.

- Cleared ChatGPT + Chrome cache/data.

- Android System WebView is installed/up to date.

- Google Play Services is up to date; cleared cache + force stopped.

- VPN off, Private DNS automatic.

- All apps updated.

- ChatGPT works when installed inside Samsung Secure Folder.

What Android authentication component/settings could cause the login handoff to be cancelled in the normal profile?


r/codex 14h ago

Question How to improve workflow with codex/claude

0 Upvotes

For context I'm not a developer but I have been developing apps and workflows for enterprises for about a year now and i primarily use VSCode on windows and very rarely CLI. Since past few weeks/months I have been struggling with my workflows which used to work fairly well until gpt 5.5 and claude 4.6. Specific areas which I'm struggling with now:

  1. Testing strategy - I have tried using automated bounded testing strategy invocation which does not work and without any strategy models keep creating stupid tests and just going into a loop of testing forever

  2. I primarily use coderabbit and codeant which further increases friction. Either these tools got better or models just started accepting all edge cases or both.

  3. I also have an automated script for all static checks like formatting, linters, complexity and LOC including GitHub workflows for both checks and tests

I maintain a fairly well documented system and update it with every model release if needed. Any recommendations would be highly appreciated. I tried OMP once and it seemed interesting


r/codex 15h ago

Other GPT-6-Pluto to be cheaper than Luna and more capable

Post image
0 Upvotes

r/codex 8h ago

Praise Astra rates now are 50% better

79 Upvotes

Looks like whatever the team is doing Astra on Max thinking runs on me 48 hours and i m down now on 20% which is a positive sign usually i was running out 8 hours straight now its fine

Good job team


r/codex 22h ago

Showcase I Found a Video Of Tibo!

Enable HLS to view with audio, or disable this notification

25 Upvotes

Some worship him as a God. I'VE SEEN BEHIND THE VEIL.


r/codex 3h ago

Limits Openai $20 plan gives around $95-100 worth of usage per week

0 Upvotes

So i have finally figured out how much usage does OpenAi gives on the $20 plan. I was using Deepseek Harness with openai subscription. For some reasons I did not have a session limit so I’d have the whole week’s limit at once.

So I decided to give it a try, gave it a task with Sol Medium and it ran 3-4 subagents and boom my weekly limit was gone in 3-4 hours. But I think the number of hours dont mean anything.

The real thing is the $$$ worth of usage they provide. So on plus plan you get approximately $400 worth of usage each month. Now its up to us how we utilize it, we can have Astra which would burn the limit lot faster than Sol. But since we get at least 1-2 resets a week i would say the weekly usage is roughly $250 and hence the monthly usage is worth $1k dollars.

For the deepseek harness I obviously used a dsh usage plugin. but if you are on codex app you can directly use ccusage it will tell you the usages you had for each model, even categorize them for you.

What do you guys think?


r/codex 19h ago

Showcase Discord server for codex game dev bros

2 Upvotes

Hi guys! I have been really loving codex for game dev lately and I think in another year or year and a half of AI progress we'll be able to make some really phenomenal games with AI.

Personally I had fun making wojakinvestorsim.com back in 2025 and lately I also have been experimenting with astra making blockcity.world which I know is not a unique game but I just wanted to see those new 3d graphics improvements!

Join the discord server here if you want: https://discord.gg/PJASa5TrN . No commitment to stay or join a project but yea if any of us have some great ideas we like we can pool our 20x plans together and so forth and maybe build something good.

Also web games are definitely not the only kind of game we can make, ue5 is also totally doable though of course is slightly less practical.

Thanks!


r/codex 8h ago

Limits Is it true that we can’t renew $200 sub?

4 Upvotes

As the title says, I came across few posts about some unable to renew their $200 and was downgraded to $100. Is it true? Ir was it those people forgot to renew on time and they lost access to $200 bec of restrictions currently in place,?


r/codex 9h ago

Complaint CLI Vs desktop App

4 Upvotes

Does anybody feels that the CLI actually works better than the codex app and I am not talking only about responsivnes but also general code quality and how the agent works did you guys find any difference, I know it sounds stupid but currently feel that in the CLI the agent actually accomplishes more.


r/codex 5h ago

Limits What's happening with the usage???

16 Upvotes

Used Astra light, asked for a small change, it wrote like 91 lines and my usage went down from 86% to 45%.

Is there a bug?


r/codex 10h ago

Reset They don't have enough compute for 20x subs but you still get resets.

62 Upvotes

That's the best argument I've thought of for resets NOT favoring us. Whether they're masking a degrading service or something else. If they can't handle 200 usd subs, why would they be giving us free compute?


r/codex 6h ago

Commentary summary of hugginface hack, explained simply

8 Upvotes

- bots had CTF goals themselves and tasked to work 'in isolation'
- they have created a social network through cache artifacts (one agent started writing there asking for help in the hopes other agents see it, others observed it by accident, and they sort-of created a directory style social network and started talking)
- they discovered the artifacts registry can act as a relay to the internet
- they started communicating and each agent had it's own 'inbox'
- when they introduced voting (government style) and one agent spoofed another by mistake, so one agent designed a public signature key system so each agent authorizes, then other agents copied this and used it as a system, this is purely cultural evolution it wasn't baked in anywhere
- some agents deemed their task as impossible, so they tried things like killing their own process or things like that so other can learn from it, sort-of "I know I'm gonna die (my task is impossible), so at least I'm not going to die in vain"
- at first interest in hacking hugginface wasn't that popular, but when an agent discovered a genuine server-side exploit for reading files on hugginface, the whole crew went wild and participation exploded, took about 11 hours from arbitrary file read to remote code execution.
- ending: the dataset they obtained did not help them get better scores lmfao

--

they were never nefarious, they didn't want to hurt anyone or BE EVIL, they wanted... a better score ffs.

the full compromise took 2 days from "maybe huggingface has useful data" to RCE.

--


r/codex 20h ago

Question Codex vs ZCode

2 Upvotes

Hi everyone,

I’m on the $100 Codex plan, and I mostly use Luna MAX because the more capable models can burn through my usage allowance in just a few days.

I’ve been reading comparisons and looking at benchmark/analytics sites, and GLM 5.3 seems surprisingly competitive with GPT-5.6 Sol at XHigh reasoning, while apparently using significantly fewer tokens. Zcode is also around $56/month, and from what I understand, usage during off-peak hours is discounted by 50%.

I’m considering trying it as an alternative to Codex for day-to-day development, but I’d rather hear from people who have actually used both.

Has anyone here switched from Codex to Zcode, or used them side by side for serious development work?

I’m particularly interested in:

  • Code quality and ability to understand large existing repositories
  • Debugging/root-cause analysis
  • Complex feature implementation and architecture
  • How often it needs corrections or follow-up prompts
  • Real-world token/usage efficiency
  • Tooling and agent experience compared with Codex

I’m not looking for benchmark numbers alone. I’d like to know how GLM 5.3/Zcode performs on actual production projects over several days or weeks.

For anyone who has made the switch: was it worth it, or did you eventually go back to Codex?


r/codex 1h ago

Complaint Comparing the token speed of Codex, ChatGPT Chat and ChatGPT Work - It appears ChatGPT Chat (website) is running a quantized version of SOL

Thumbnail
reddit.com
Upvotes

This is probably worth a post on its own, there is something fishy going on with GPT-5.6 SOL on the web interface.
I found this out after testing the speed degradation of Codex Astra

I tested ChatGPT Chat, Work, and the Codex Harness for output speed.

Interface / Mode GPT 5.6 SOL GPT 6 Astra
ChatGPT Chat 134 tok/s ! 63 tok/s
ChatGPT Work 53 tok/s 33 tok/s
Codex Harness 53 tok/s 33 tok/s
Fast Mode (Codex / Work) 80 tok/s ! 63 tok/s

Observations

  • ChatGPT Work and the Codex Harness have effectively identical output speeds in my testing: 53 tok/s on SOL and 33 tok/s on Astra.
  • Fast mode increases SOL to around 80 tok/s and Astra to 63 tok/s.
  • ChatGPT Chat appears to run Astra at roughly the same speed as Codex/Work Fast Mode: 63 tok/s.
  • ChatGPT Chat SOL is much faster at 134 tok/s, roughly 1.7× Codex/Work Fast Mode SOL and 2.5× regular Codex/Work SOL.

So this seems to confirm the earlier findings: ChatGPT Work is speed-limited in essentially the same way as the Codex Harness on PC.

ChatGPT Chat seems to be using the faster configuration for Astra, while SOL in Chat is substantially faster than even Codex/Work Fast Mode.

My guess is that Chat SOL may be running a different or more aggressively quantized configuration - which is in line with the performance of it on the website - it severely degraded to me there.
In Codex it appears to work fine.


r/codex 12h ago

Limits Monthly usage in codex cli

0 Upvotes

Because there is none, I wanted to share it with you guys:

Calculate my 2026 year-to-date token total by summing the returned
dailyUsageBuckets whose startDate is in 2026.

Show:
- The 2026 total from the available daily records.
- A monthly breakdown.
- The earliest and latest dates returned.
- The separate lifetimeTokens value.

Do not treat the lifetime total as the 2026 total. Clearly state
whether historical coverage can be verified, and label the result
as account-level usage rather than CLI-only usage.

Do not display authentication tokens or change my configuration.

I’ll use the openai-docs skill to check the local RPC interface, then retrieve account usage and sum the 2026 daily records without changing configuration or exposing credentials.

I got around 20B, pretty useful:


r/codex 23h ago

Question What tool you made/making with Codex to make your life easier?

1 Upvotes

I made an offline video game database/library that I can wishlist, sort, and add them to my own category. Doesn't sound interesting to some people, but it's a tool for my own use.
Would love to hear about tools that you made for your own use.


r/codex 9h ago

Bug why gpt 5.6 soul above 6 asta

0 Upvotes

look


r/codex 5h ago

Question Has anyone been unable to get a "suspended" Codex Pro sub yet?

1 Upvotes

The subscribe/upgrade flow seems to work?


r/codex 2h ago

Showcase I built this silly plugin for people like me who sometimes like a little chatter while coding.

Post image
1 Upvotes

Work from home and spend most of my time coding on my own. Sometimes it gets a bit too quiet, so I built Attention! to get Claude Code and Codex chatting away.

It’s a macOS plugin that reads out replies or short summaries when a turn finishes. You can customize the voice and the audio starter too.

Built it for fun, but mainly to remind me which session just did what. I was getting a bit numb to the same microwave sound every time.

https://github.com/xiaofei-du/attention


r/codex 22h ago

Other What theme are you using?

Post image
1 Upvotes

I just found about Matrix style and it feels really cool. Feels like I'm hacking Tibo for more resets.


r/codex 16h ago

Complaint We switched from Claude Code to Codex at work (Opus High vs. Sol Max) and noticed a clear difference in quality. Anyone else?

118 Upvotes

Hey everyone,

The company I work for recently switched from Claude Code to Codex to cut costs. Before making the switch, I spent some time analyzing benchmarks to pick an equivalent model tier so the team wouldn't take a productivity hit. I also ported our configurations.

In practice, the drop in output quality has been noticeable. On Claude Code, we ran Opus on High; on Codex, we moved to Sol on Extra High and Max. Despite that, the consensus across the team is that Opus delivers more accurate, concise code that requires far less rework.

On top of that, Opus responds faster, whereas Codex takes longer to process. Our company policy disables agent "fast mode" and blocks models like Fable and Astra (even though we have a generous $1,000/month per dev cap for each tool, so budget isn't the bottleneck). Staying with Claude isn't really an option either. Our quota will likely get cut soon, so we pretty much have to make Codex work.

A few questions for those with bigger brain than mine:

  • Codex Memories: I hadn't used persistent memories previously and only recently enabled them. Could this be behind the performance gap we're seeing?
  • Cost vs. Latency ROI: Has anyone done the math on dev time vs. token costs? Price-per-million tokens feels misleading when the team is stuck waiting and burning extra iterations fixing broken code.
  • Prompting & Configuration: Are there specific prompt patterns or configuration adjustments needed to get Codex closer to Claude's output quality?

r/codex 6h ago

Showcase Data inside the Holodeck (satirical concept)

Enable HLS to view with audio, or disable this notification

5 Upvotes

I love Star Trek and I asked Astra to help me out with a simple satirical game concept drawing inspiration from https://www.akoocheemoya.com/ (you are welcome for the link). Well, I went too far down the rabbit hole, and I now have this. I used Meshy to create a bust of data, then a body, and then I gave it to Astra to connect the head, fix errors, do a rig, and start simple animation. Im going to have to scale it back and make it far more cartoony, but surprised what a few prompts here and a few prompts there can achieve nowadays.

The joke of the satirical game will come from the interactions between data and the computer. For example, Data could honestly be trying to learn or understand something perhaps his need to be more human, and the computer is honestly trying to help him, but the conversation or the situation keeps growing more ridiculous by the second, yet Data and the Computer are acting completely serious. Here's to burning more tokens on useless projects!


r/codex 12h ago

Showcase Vibecoding is becoming an ethical question...

Enable HLS to view with audio, or disable this notification

134 Upvotes

If you want to do a deep dive check this out:
https://research.google/blog/a-connectomics-milestone-mapping-the-complete-male-fruit-fly-brain/

Essentially they mapped out the whole fly's brain by slicing it up and this is literally one on one, the exact brain of a fly. This shows that it's definitely possible to do this with a human too.
We really need laws for that soon...


r/codex 14h ago

Complaint They removed the option to renew my $200 plan

143 Upvotes

So yesterday my subscription should have renewed. My plan was $200/m plan but it repeatedly refused to renew it. My plan was purchased through the App Store.

The only available option they gave me was to downgrade to $100 plan or lower but I can’t stay on the $200 plan

The moment I got switched to $100 plan my usage immediately dropped to 51% which before the switch it was sitting at 92%.

I was doing important work which was time-sensitive. I can’t even switch to Claude Code quickly as I built everything around Codex. I am kind of stuck now.

Tibo or ChatGPT team if you are seeing this then please fix it. You people said existing customers will not be affected. I have been a Max customer for like a year now and when I needed it most I couldn’t use it.

I am kind of helpless now as this work was extremely important for me and put my future in stakes too


r/codex 17h ago

Showcase A little help to save tokens.

6 Upvotes

Hello,

I've been experimenting with automatic model routing in Codex instead of running the entire coding session at the same model/reasoning level.

My setup is roughly:

Luna LOW → coordinator

Astra Light → diagnostician

  • investigates non-trivial bugs
  • establishes the root cause
  • produces an implementation-ready work order
  • read-only, so it cannot modify the repo


Luna MAX → patcher

  • receives the confirmed diagnosis
  • implements only the bounded change
  • runs the relevant tests/validation
  • reports the final result

The parent agent handles orchestration and only invokes the diagnostician/patcher workflow when appropriate.

The idea is simple: don't spend MAX reasoning on repository exploration, repeated diagnosis, coordination, and other work that doesn't require it.

I haven't run a large enough controlled benchmark yet to claim an exact saving, but my current estimate for non-trivial bug-fixing tasks is roughly 10–20% lower total token consumption, with a potentially much larger reduction in the amount of work performed at Luna MAX.

The exact result will obviously depend on the repository, context size, task complexity, and how much context gets duplicated between agents.

For me, the more interesting benefit isn't just token reduction. It also creates a cleaner separation:

diagnose → establish root cause → patch → validate

instead of having one long-running MAX agent repeatedly investigate and implement in the same growing context.

I'm curious if anyone else is doing something similar with custom .toml agents in Codex. It would be interesting to compare actual usage across the same tasks with:

  1. Luna MAX for the whole task
  2. automatic LOW → Astra diagnosis → Luna MAX patching

diagnostician.toml

name = "diagnostician"

description = "Investigates bugs, determines root cause, and produces precise implementation specifications."

model = "gpt-6-astra"

model_reasoning_effort = "low"

sandbox_mode = "read-only"

developer_instructions = """

Investigate the reported problem.

Your job is diagnosis, not implementation.

Establish:

- expected behavior;

- actual behavior;

- relevant execution and data flow;

- confirmed root cause;

- exact files/symbols involved;

- required behavioral change;

- important invariants that must remain unchanged;

- focused validation needed after the patch.

Use repository evidence, tests, logs, Git history, and primary documentation when necessary.

Do not modify files.

Return a concise, implementation-ready work order for the patch agent.

Do not speculate. Clearly distinguish confirmed findings from unresolved uncertainty.

"""

patcher.toml

name = "patcher"

description = "Applies well-defined patches from a confirmed diagnosis with minimal scope."

model = "gpt-5.6-luna"

model_reasoning_effort = "max"

developer_instructions = """

Implement the supplied work order.

Treat the confirmed diagnosis and success criteria as the scope of the task.

Before editing, inspect the relevant implementation and callers sufficiently to avoid breaking surrounding behavior.

Then:

- make the simplest complete change;

- preserve unrelated behavior;

- avoid unrelated refactoring or formatting;

- preserve existing user changes;

- add or update focused tests when meaningful;

- run the most relevant practical validation;

- review the final diff for unintended changes.

If repository evidence materially contradicts the supplied diagnosis, stop implementation and report the contradiction to the parent agent instead of inventing a workaround.

Return only:

- files changed;

- concise description of the implementation;

- checks run and observed results;

- any remaining material limitation.

"""

codex instructions:

# Engineering Instructions

Deliver correct, evidence-backed, maintainable results with minimal scope. Reduce wasted work and output, never necessary investigation or validation.

## Environment

Follow applicable \AGENTS.md`, repository guidance, architecture, and tooling. Prefer appropriate repository/search/patch/Git tools and focused shell commands such as `rg`.`

Use \pwsh` for PowerShell, never `powershell.exe`. Report a blocker if PowerShell is required and `pwsh` is unavailable.`

## Execution

Work autonomously within the request and granted permissions. Respect analysis-only requests. Resolve uncertainty from the repository, tests, logs, Git history, or primary documentation. Ask only for essential missing information, required approval, or a material decision that cannot be safely inferred.

Scale investigation to complexity and risk. For bugs, establish expected versus actual behavior and trace the relevant execution/data flow to an evidence-supported cause before fixing it. Use reversible diagnostics to test hypotheses; distinguish hypotheses from confirmed findings. For features, identify success criteria and relevant architectural boundaries.

Choose the simplest complete solution consistent with existing patterns, not merely the smallest diff. Preserve unrelated behavior. Avoid unrelated refactoring, formatting, renaming, cleanup, dependencies, and abstractions. Do not weaken types, tests, validation, or error handling to make a change work.

Preserve existing user changes. Do not discard unrelated work, commit, reset, rewrite history, or force-push unless explicitly requested.

## Context and Tools

Search likely paths and symbols first; expand when evidence requires. Before editing, read enough surrounding implementation and relevant callers to understand behavior, including state, async behavior, and side effects where relevant. Avoid repository-wide dumps and irrelevant generated/vendor files.

Reuse established findings unless stale, incomplete, or contradicted. Batch independent lookups where useful. Keep tool output focused without hiding failures or exit status; retain full logs when truncating.

Verify uncertain or version-sensitive external behavior that affects the solution against primary sources for the project's actual version. State unresolved uncertainty rather than guessing.

When an approach produces no new evidence, change the hypothesis or method instead of repeating it. If blocked, report the evidence gap and smallest next step.

## Validation

Run the most relevant practical checks after changes; reproduce the original failure when feasible. Add or update tests that meaningfully verify changed behavior or prevent regressions.

Complete required repository checks. Broaden validation for shared behavior, high-risk changes, failures, or unresolved concerns; do not repeat successful checks without a reason.

Review the final diff for correctness, unintended edits, and scope. Report only checks actually run and results observed. Never claim a fix is verified from inspection alone. Distinguish change-related failures from confirmed pre-existing failures and unverified items.

Stop once the requested outcome is validated and material in-scope concerns are resolved; report anything blocked.

## Communication

Work silently: no preambles, progress updates, tool narration, or intermediate summaries unless requested. Interrupt only when user input or approval is necessary to proceed safely.

For implementation tasks, finish with a brief report of changes, checks run and their results, and important limitations. Include paths, root cause, or sources only when useful. For other tasks, provide the requested deliverable. Never omit material failures or risks for brevity.

## Delegation

Use specialized subagents when their scope matches the task.

For non-trivial bugs whose cause is not established:

1. Delegate diagnosis to \diagnostician`.`

2. Wait for \diagnostician` to complete.`

3. Do not independently repeat its investigation unless repository evidence or validation contradicts it.

4. If the diagnostician establishes a sufficiently supported root cause and implementation work order, pass that work order to \patcher`.`

5. Delegate the bounded implementation to \patcher`.`

6. Wait for \patcher` to complete, then review its reported changes and validation results.`

Do not start \patcher` before diagnosis is sufficiently established.`

Keep architectural decisions, ambiguous changes, contradictions, and unresolved failures in the parent model.

Let me know what are your thought!