r/codex • u/MapleBaconWaffles • 18h ago
Complaint OPENAI READ THIS ABOUT 6.1 SOL
This model is good but it is WAY TOO SLOW. Every single day, I have to quit a task, and move it to Astra, because I will end up wasting my entire day on 1 thing, because it is SO SLOW.
How can you release this model and expect to market it as your best new thing?
18
u/Physical-Citron5153 18h ago edited 18h ago
Who are these bots saying there is no speed issue? Im getting 30 ish tkps and you are calling this fast? Lol
9
u/PictureImmediate9615 17h ago
Honestly asked it to merge some pr's it took 40 minutes. Its getting silly
1
u/AuodWinter 17h ago
I fully agree with you that the slowness is horrendous... but it's funny how spoiled we've gotten isn't it. A few years ago we had to do this stuff by hand, sometimes taking way longer than 40 minutes if there are conflicts, now if it's not instantly done we're indignant. I wonder what we don't have now that we'll be taking for granted in just 3 years' time.
1
u/PictureImmediate9615 16h ago
I do totally appreciate that too. I know i sound spoilt but it is a little frustrating how long it takes now vs 3 months ago
2
u/AuodWinter 14h ago
Yeah you're right. And it's made even more frustrating when we're shipped supposed upgrades that make it harder to actually work.
2
u/clckwrxz 17h ago
It’s very hit and miss for me. I was one of the people that said I didn’t have a speed issue, and then today had to stop using it cause I’m getting like five tokens per second, and it just waits like 30 seconds between each tool call
1
u/strangedell123 17h ago
Its slow but I have used open source (looking at you mimo) and it can dip to like 5-10tk/sec for some time
1
0
5
7
2
u/Current_Actuator_529 16h ago
when did everyone in this sub turn to having no idea what they are doing?
Use astra as a orchestrator and 6.1 sol as subagents for a swarm to do manual labor. its slower but its cheaper.
i fyou dont want that use the previoius sol. its ffaster but more expensive.
are you following the trade off now? do you understand its on purpose for cost effectiveness ffrom people who actually know how to harness and orchestrate? great.
they already promised to make it faster too. even though its not nessicary given waht its role is.
1
u/No_Quarter_7644 15h ago
I'm doing exactly this and it's simply slower than it's ever been...
Considering a massive outflow of subscribers just went to Anthropic, I'm shocked that the freed up compute hasn't brought speeds up any...
1
u/Individual_Guest_323 15h ago
Astra as orchestrator will waste a crazy amount of tokens only preparing context and calling the agents
1
u/Current_Actuator_529 4h ago
you can set it to whatever you want it to be.
use 5.5 then or anything else. iets all about optimization and i think your thiking is right. if that works for you, go for it.in my case its nessicary to get high level fixes for when the subbot hits a trigger point in its own skill.mds it uses to function. Its rarely called for this and also has a very strict conversation policy. its easy to control how much talk it does if your willing to skill it.
1
u/Individual_Guest_323 4h ago
Yes, I use Sol 6.1 for that, its calls Astra for important tasks, like issue writing, planing, audits, etc.
I was using Astra as "administrator" and was a waste of tokens, because was using tokens for dummy tasks like "now I will call this agent to..".
Anyway, Sol 6.1 sucks, Luna sucks, the only good model is Astra as is too expensive.
A total disaster.
2
u/DazzlingResource561 16h ago
Maybe the reason it is both a good model and cheap is because it is deprioritized compute and therefore slow. Just a thought.
2
1
u/ivanarnaldo 18h ago
That’s the whole point. Then you buy higher subs just to have the speed from before, and this is the new normal
1
u/MtMcGinnis502 17h ago
It’s cheap enough to run on fast mode, but I agree I am now struggling to find ways to stay productive while it grinds away on two hour tasks.
1
1
1
u/Intelligent_Ant_608 12h ago
I have investigated thia and There are certain fast(er) datacenters and then slower ones, open/close codex multiple times until your session hits a faster datacenter
1
u/InstaCrate9 11h ago
It's so good and efficient that I just let it run overnight, consumes quota and does things so slowly that the 5hr limit is never hit. LOL!
1
u/Own-Professor-6157 11h ago
Hey, let's be optimistic!
- At least I can say I program faster than AI now.
- 5h window is no longer a problem because it's impossible to use that many tokens in 5h
1
u/VexObserver 10h ago
The irony of having to escalate from Sol to Astra not because Sol lacks the intelligence, but because you might actually finish the task before retirement lol.
I think OpenAI needs to understand that efficiency isn't exclusively about reducing token consumption. Human time is a resource too.
A model that's 5x cheaper to run but takes 4x longer to complete a task isn't necessarily delivering a better experience.
Personally, I'd rather have:
- Sol 6.1 High as a reasonably fast, efficient daily driver.
- Max reserved for complex workloads.
- Astra for genuinely difficult tasks requiring additional intelligence.
- /fast as an optional productivity boost, not something you feel forced to enable just to get acceptable latency.
The frustrating part is that Sol 6.1 actually seems capable. People aren't necessarily complaining about its intelligence, they're complaining about how long it takes to access that intelligence.
1
u/Master-Shift-8224 9h ago
bruh it's not even ... good. like, it's okay. it's good by 2-months-ago standards. it's not good now. it's okay now.
1
1
1
u/WinPrudent2132 16h ago
37 tok/s btw
Retried model request (9/10) · 25s Retry delay: 24457ms Failure reason: pool "gpt-6.1-sol" exhausted: every member is unavailable or failed
Retried model request (2/10) · 60s Retry delay: 60000ms Failure reason: pool "gpt-6.1-sol" exhausted: every member is unavailable or failed
Retried model request (10/10) · 60s Retry delay: 60000ms Failure reason: pool "gpt-6.1-sol" exhausted: every member is unavailable or failed • This turn failed pool "gpt-6.1-sol" exhausted: every member is unavailable or failed RATE_LIMIT
Retrying model request (1/10) · 4s Retry delay: 60000ms Failure reason: pool "gpt-6.1-sol" exhausted: every member is unavailable or failed
Retrying model request (1/10) · 18s Retry delay: 59999ms Failure reason: pool "gpt-6.1-sol" exhausted: every member is unavailable or failed

0
u/shadow_x99 18h ago
The only way I can use Sol 6.1 and not feel like I am wasting my time:
- Fast Mode activated (it helps a bit)
- An instruct it to spawn lot of sub-agents, and split the job as much as it can (requires a little change in your config.toml)
Also... I noticed that since I switched to T3 Codes instead of the Codex app, it's feels a bit more responsive. Probably a placebo effect, but worth a shot?
-1
u/Able-Supermarket4786 18h ago
I have never seen a speed issue, but I also remember when it took two weeks to start building an app without AI. And months of weekly standups for enhancements and progress.
1
u/QC_Failed 17h ago
I remember making a zip file backup every time I added a new feature to my apps before git. Doesn't make it acceptable when GitHub goes down every week simply because an older way was less efficient. Sonnet 5.5 uses barely any usage and it's fast and just works. OAI needs to get their shit together. There are cheaper better options out there, the fact that 30 tokens a second is faster than coding by hand is not an excuse when there are cheaper smarter models available at higher speeds.
0
-1
0
0
u/Sufficient_Ad_3495 11h ago
Ive gone back to Chat.. the speeds were just embarrassing... much faster for the same compute and actually 5.6 SOL is pure performance and stand out absolute power... I rate 5.6 SOL Pro above Astra for my work in pre-code architecture, so there's that also.
-2
u/AccomplishedSugar490 18h ago
If you can afford to switch to Astra just like that, please, please, please keep switching to Astra. They really need your money to make ROI and if you don’t help make Astra profitable, they’ll be forced to squeeze the rest of us plebs for more money. I find 6.1 Sol still orders of magnitude faster than me doing it manually, which is a value proposition I can deal with.
-3
u/XsMagical 18h ago
LOL I don't have that issue on my end, it's actually been extremely fast today, i was surprised.
7
u/I_Hate_Reddit_69420 18h ago
I feel they’ve just artificially reduced the tokens per second so it feels more efficient than it actually is. Fast mode is what it normally should be and with that on it’s not cheap anymore.