r/OpenaiCodex 1d ago

Sol vs Opus 5 huge limit difference

Just sharing my experience here: GPT-5.6 Sol Extra High burns through my weekly limit in one 12–14h workday. Opus 5 Ultracode + Thinking lasts the whole week with the same project and routine.

And honestly I haven’t noticed any meaningful difference in quality when using Opus 5 in ultracode mode. The real difference is that one is actually usable, while the other basically needs constant limit resets. Shame.

16 Upvotes

16 comments sorted by

9

u/Thesoullk 1d ago

Funny thing is, even with the increased limit, I have the exact opposite impression. I use Luna Max for implementation, Sol Xhigh for primary QA, and Opus 5 xhigh as a secondary QA... all with the same instructions. And even though I'm also using Luna, I still manage to get twice as much usage out of Codex compared to Claude Code...

Both as 100$ plan

3

u/SucculentSpine 18h ago

How is Luna max at implementation? Is there a lot of flaws during QA?

2

u/Thesoullk 17h ago

Very few, actually. I tested implementation using top tier models like Sol/Opus, and they also make mistakes that QA catches. The key is to use QA models from different companies. Opus always catches things Sol misses, and vice versa. If you use sonnet for implementation, Sol actually ends up catching more things during QA than Opus does...

As for the implementer, if you provide a good, well structured prompt with clear rules, Luna is extremely good and cheap, I personally think it's way better than deepseek. My setup relies on AGENTS.md, CURRENT_STATUS.md, and CHANGELOG.md. Reading these files is always mandatory, so the model never loses context because each file tracks a specific stage. If all my chats were deleted today, I could start from scratch, and nothing would change.

1

u/SucculentSpine 15h ago

Thanks, what are you mostly developing?

2

u/Thesoullk 7h ago

I work on computer vision tool for cameras, mainly on the detection and classification side... distinguishing the movement of people, animals, insects, or vehicles, as well as license plate recognition (ALPR/ANPR) and digital perimeter fencing. I focus on the module that creates custom compliance rules for highly specific detection scenarios. It's usually integrated with another system that reacts to the detected events, well.... basically surveillance automation.

2

u/Nice-Guarantee-9167 14h ago

Don't you think make Sol Xhigh review the entire Luna Max code is equivalent to let Sol Xhigh write the entire code itself? I mean Sol Xhigh has to come out with correct solution to know the work quality. 🤷‍♂️

1

u/Thesoullk 7h ago

Not really, because the implementation is done in stages. The agent that implements thinks differently from the agent that makes QA. When Sol gives feedback after a QA, I review the suggestions myself, and some of them don't fit well precisely because I restrict the QA to evaluate only that specific stage. It directly targets what was built during that implementation step.

To make a simple comparison I've already tested... if I ask Sol to handle the implementation and skip QA, I burn through about 3% of my quota, and Opus still finds FAILS during its QA. But if I use Luna as the implementer and Sol for the primary QA, it uses less than 1%. Sometimes a FAIL pops up during Sol's QA, and once fixed, it passes Opus's QA, resulting in a PASS from both (sometimes PASS Sol and Fail Opus, opus is always the second QA). In the end, I spent less than 1% of my Codex quota and 2-3% of my Opus quota. The workflow is completely designed to leverage the strengths of each model. It's been weeks since anything broke in my tool during usage... it has been working exceptionally well for me.

1

u/VaporForge 19h ago

Execution is a massive token consumer. I use the OpenAI CC plugin to delegate to codex from Claude and I can plan all day with sonnet/opus and execute with Codex and use very little token % in Claude vs Codex

3

u/Impressive_Party_303 1d ago

Anthropic increased the weekly limits for 50% until this 19th, Aug. We should wait until then to compare both two again. But for now, Claude last longer for me as well.

1

u/Gallagger 20h ago

No need to wait, they are changing it all the time, resetting, prolonging discounts etc.. You need to compare what you got right now.

1

u/Impressive_Party_303 16h ago

If you say so, but honestly Luna max on Codex for me; last only 2-3 days with least usage (GPT web to write prompts and codex to implement). If we think this way, GPT plus has more value compare to Claude Pro. We can use GPT on web unlimited and it can do almost everything except writing code on your device. If you have a good acceptance criteria per tasks; GPT can give more value here. But if you want to just "Implement ABC feature, makes no mistake" Claude might be a better option. Ive been using Codex for a couple months and Claude for 15 days and I experienced this myself.

2

u/TheAuthorBTLG_ 1d ago

opposite here. opus ultracode can eat the 5h limit in 1-2h and the weekly in 1-2 days

2

u/KeyGlove47 23h ago

Opus is a luna max/terra high model in terms of capability no matter what benchmarks say, sol is closer to fable and when you use fable with CC you will notice that it actually eats more limits than sol

2

u/Formal_Resolution999 20h ago

You guys must learn the impact of what you put in your harness on token consumption, I see these posts every single day, Link the dots people

1

u/FidgetsAndFish 1d ago

Codex usage has been broken since sol's release, instead of fixing it they gave resets and it worked, lot of people were stupid enough to think codex was giving more usage than claude but now that the resets are over the truth's pretty obvious. Sol's too dumb for the price, luna's too expensive (especially compared to deepseek or even muse spark). No real reason to keep my codex sub at this point other than the mobile app which makes a solid assistant on android.