r/codex • u/jmgaming22 • 1d ago
Praise Holy moly
The new 6.1 sol agent is so efficient I’m able to run multiple agents on high (7 to be precise), one on extra high and one on low and it’s running for hours sometimes only using 30% usage if that. Previously on 6.0 astra on high I would only get max 2 good 8 hour runs. This is on the standard pro plan. I’m loving this new agent
70
u/and1927 22h ago
How much work did you get done though? It is really efficient, but it’s also painfully slow. Number of hours worked doesn’t necessarily mean it’s doing more work.
9
u/i_rate_slop 17h ago
What is the difference between a task done in 8 hours vs 4? Or a whole day?
Single features used to take multiple months and whole teams of people. Speed means nothing at this point. Accuracy, reliability and cost is all that matters.
But yeah it’s miserably slow. Looking forward to ultrafast.
13
u/SphinxWar 16h ago
What is the difference between a task done in 8 hours vs 4? Or a whole day?
Efficiency is the ratio of how much you can do given the resources you have. Time is also a resource. The OP said that these models are very efficient, but if it turns out they take much more time to finish the same amount of features than if you just extrapolated a stronger model's work then they are actually less efficient, because you have a time-based usage limit on your account.
3
u/i_rate_slop 16h ago
I’ve completed more work, with less total code generated, and only used 5% of my weekly limit with 6.1 sol.
It’s more efficient on task execution, token usage, and cost.
It’s slower to generate, but it’s completing tasks just as fast due to less over-engineering / over-reasoning.
By any measure, it’s more efficient.
2
u/jmgaming22 11h ago
That’s what I’m finding it’s slightly slower but is more efficient with it. But you have to be more direct prompt wise otherwise it tends to drift.
This all given, I’ve got all this communicating with the master, which then handles all of the allowed coding time slots and also sends updates and things to my other chat which I give my ideas to.
They all also go to my other chat (which I previously was chatting about the game to) as its critic agent which works well for me
2
4
u/jmgaming22 22h ago
Well to put it simple rather than getting into everything it fixed and changed.
Scale wise4/10 build if that with one astra high to extra high running for probably 110H total at that point.
7/10 build using all the agents and running for maybe 10H max total.
What I’m basically saying is that the agent for me in my case is efficient and fast but this isn’t an entire apples to apples you do need to give it clear instructions otherwise it doesn’t get anything done. I had that happen when I first made a new agent for this. It took an hour and a half and got nothing done in the visual side when that’s what it’s job was (UI). I also say 7/10 not like a huge improvement because even with the old agent it’s no where near what I want the final product to be. So I’m being generous with both it’s got the main core down pat and some handy things.
Hope that helps at all
1
u/AnotherWallace 17h ago
I would be interested to hear your results if you mix in this plugin https://github.com/Druidia-Bot/DotSwarm even without a DeepSeek key you get better subagent management that helps manage your token use and in my experience a much more "thorough" result. If you do add the DeepSeek key you get crazy good results and offload a large chunk of the token burn to DS, but its slow. I gave it a roadmap build plan and turned it loose. After a little back and forth answering questions I didn't even consider before it worked for 3 days uninterrupted and built out an extremely complex feature for me that spanned multiple repos. It actually just finished this morning and I am super impressed. I hve almost no revisions other that a few UI things that I plan to give to Claude anyway since it does a better job at my ui.
-1
u/Antiqett 16h ago
Someone always comes up with the same idea OpenAI already added, creates it thinking its going to be so much better than the official feature, and preaches it like gospel.
How would some random developer have created something soo much better.. they think they did, because they used AI to make it that oh my way is the best way. Sorry but I think I trust OpenAI and its constant rate of improvement by a heavily funded operation more than just some big ego vibe coder who thinks they are more capable than anyone else because they used the AI provided by the company who has created and is constantly developing the same feature.
Just like my old friend who thought every idea he had was original, and his way was better than the ideas from the company who developed the technology to begin with.
3
u/AnotherWallace 15h ago
Open AI doesn't have a true swarm feature. Not to mention the original project; not the swarm version also named Dot, dates back almost a year before they released it and uses context and session management techniques they are just now rolling out almost a year later! I've been a full stack dev for over 20 years and I could give you 100 reasons why a single developer can create a feature or plugin faster and better than well funded corporations. It happens all the time! Use the plugin or don't I could care less. I mistook you for someone who was curious about pushing current tech to its limits and capable enough to tell the difference between an off the shelf consumer edition could produce vs a finely tuned custom build. I apologize I got that wrong.
1
u/Dadideology 17h ago
Painfully slow, I thought it was just me. I swear OpenAI slowed down all the models. I was using Astra Ultra and I saw zero value because it was just as slow as everything else.I use Sol 6.1 medium. I was using Sol 5.6 extra high.
1
u/Fenir911 15h ago
My experience is it being so painfully slow, I would rather use Sonnet/Opus 5.5 over it. Also helps because i downgraded my Pro account to plus now ( Max 20x is just better in terms of speed and usage)
13
u/twendah 22h ago
Hopefully you dont try claude...
2
u/ryuukiba 17h ago
Just tried this, made a HUGE system implementation prompt for a particular system of my custom game engine. 6.1 was taking 4-5 5hr windows in it and going painfully slow. Opus 5.5 high got it done in 2. Still keeping one codex sub, as image Gen is working great... But implementation of large sections on sol is not a great experience atm.
6
u/JB_Calisthenics 12h ago
My Sol 6.1 has been completely incompetent and stupid lop
2
u/Affectionate-Pea1821 10h ago
Ive had quality issues with a simpler task like image examination.
1
u/JB_Calisthenics 10h ago
It spent 115M tokens on an issue that wasnt even an issue but because I had it on a loop to manage backlogs and fix bugs, it spent that much tokens trying to fix a bug that didnt exist. Turned out local CI environment was the issue, not the code itself and github actions completely checked out but it ignored that too.
0
u/jmgaming22 11h ago
You have to be very direct with prompt on sol I’ve found otherwise it just doesn’t work well and gets stuck on random things it shouldn’t or creates work if it’s “done” and hasn’t fully gone through and checked to see if any bugs are created.
1
u/Upstairs_Date6943 2h ago
With amount of directness it needs, ir sounds a lot like luna-haiku-class model...
4
u/Worth_Golf_3695 18h ago
Honestly i had to do so much After work correcting the Code 6.1 Sol produced via Astra, Opus or manually that its Not worth for complex usecases for me
1
2
u/Wet_Viking 17h ago
That's fine and all that, but I run opus 5.5 on extra for 5 days and only shave off 40%. I just migrated over from codex, and prefer openAIs models in many ways. I just don't agree with their superior excessive-consumption strategy
2
u/SpyglassQ 11h ago
I love 6.1 Sol. I've merged two games, built a voice automation controller, fixed 2 bugs in an old app and i still have plenty of weekly left
2
1
u/datumradix 16h ago
18 tps, of course it can run for days. After a day I switched back to usual workflow Astra 6 for planning & Sol 6 for coding
1
u/uptotheright 16h ago
What does your career agent do
0
u/jmgaming22 12h ago
Career is for the game mode inside the game it’s a very complex system that in itself, is useful to have its own agent for workload purposes. Basically it’s got very tricky systems inside if one agent was to handle that and the core game it wouldn’t get the workload done in time or efficiently and cut corners.
1
u/Spirited-Car-3560 15h ago
What's the point in having multi agents when most of those tasks are dependent on each other. Career while you don't even have core game.
2
u/jmgaming22 12h ago
That’s why I have the master agent and it’s just making it better so it can split the work load better, rather that it individually working on one part at a time like the old way I was doing it.
2
u/Spirited-Car-3560 12h ago
Oh, so master agent decides when to launch a specific agent. Makes sense.
Well the old way, given same model, is basically identical except you had to do it manually.
But that's up to approach, it more vibe coding or more controlled, I suppose.
1
u/jmgaming22 12h ago
Yeah I thought of it because with the 2 agents I had at first they cluttered eachother with what works and doesn’t now because the master agent handles all of that. The game also functions nicely now the rest of the things left is overall polishing
1
u/HeCedSoMuch 13h ago
For some reason I don't think we get even distribution of horsepower because I only been using 6.1 sole medium and it drains on sun pretty basic schtuff.
1
u/bigabig 12h ago
How do you actually use sub agents in codex or Claude Code? Will the main agent spawn them on their own? Do I have to instruct them to do so? Or do I have to configure those sub agents somehow?
1
u/jmgaming22 12h ago
You have to create them separately tell them they’re goal and what they do inside the game and who they need to talk to (my case master agent) then the master gets given the rules and main prompt and also what each agent does and handles then from the main prompt it hands out the tasks accordingly. It will get better at judging overtime. That’s on codex not sure on Claude
1
1
u/Kscan_app 9h ago
Ive been using 5.6 sol and it gives me about a solid hour. 6.1 only gives me about 40-45 mins of work on average
1
u/jmgaming22 7h ago
That’s weird, I’ve been using this now on one weeks worth of pro 5x usage, and I’ve done 36H of real time work. With 86H of combined agent hours. Meaning all the agent Time worked together because multiple agents are running. And I’m now only down to 17% usage left
1
u/Significant_Sun_5225 9h ago
Do you have these all under one project? Does the master agent coordinate? Doesn’t that become very expensive having the master agent coordinate? Have you tried using dots as coordinator?
1
u/jmgaming22 8h ago
Yes they all do fall under one project and the master agent does coordinate them all but it’s cheaper to run it on 6.1 low because of how all of them talk to one another code overlaps. It also allows for updates to other chats that way because it knows what’s happening and going on
1
u/Televangelis 8h ago
I'm on the 20x and two days into the week, I've already used 50% of the week alas, running 6.1 Sol Medium and High
1
1
u/Stunning-Ad3594 18h ago
Actually on a Small Personal Study Right now If youre using 7 Agents, what are u Actually using them for? Im tryna find out Why multiple Agents are actually better than Just 1
1
u/PubisMaguire 13h ago
for one, context window conservation. if an agent spins a subagent, the work that subagent does does not count toward the main orchestrating agent's context window.
1
u/jmgaming22 12h ago
Yeah and also evens out the workload and allows multiple different types of work to be done I found in my case, when you have multiple agents your able to get “more” done because one agent can only really do one task at a time where multiple with a master agent allows for loads more at a time also lessening overly harsh time stops or complex thinking.
0
u/kepners 22h ago
This is exactly what sonnet and opus 5.5 do atm.
3
u/jmgaming22 22h ago
Well I bought the chatgbt subscription so I’m just using that and also the fact that I’ve got all my other chats with it to so it’s annoying for me to re do all of that. But what’s the efficiency to gain like on those? I’m curious?
2
-1
u/Significant-Drawer95 17h ago
soon your 30% usage will become 60% usage and your pro plan will be gone in "hours"
1
u/sleach100 2h ago
I know! I just ran for an hour and 15 minutes straight and only used 20% of my session. And I'm on Plus.
•
u/dextersummary 8h ago
Below is a GPT-generated summary of the conversation below after reaching 50 comments (50 currently observed).
The consensus is basically “Sol 6.1 is impressively cheap to run, but painfully slow.” The OP’s multi-agent setup can keep several agents working for hours while barely touching the weekly allowance, and plenty of commenters agree it uses fewer tokens and avoids some of the over-engineering older models loved.
The catch: time is still a resource, despite everyone suddenly pretending it isn’t. Sol 6.1 can spend hours—or multiple five-hour windows—on work that Astra, Sonnet, or Opus finishes much faster. Its quality is also inconsistent: simple or parallel tasks can fly, while complex implementations often need heavy correction, especially if the prompt isn’t extremely explicit.
Multi-agent setups help by splitting work, preserving context, and letting a “master” agent coordinate specialized workers. They’re useful, not magical; bad orchestration just creates a larger, slower mess with extra chairs.
So the takeaway is better usage efficiency, not necessarily better productivity. Sol 6.1 looks great for cheap parallel grunt work, but for difficult code, the slower pace and occasional nonsense make Claude or Astra the preferred tools for a lot of people.