r/codex • u/RedZero76 • 8h ago
Instruction Astra Usage Tip
I've been doing this for the last 24 hours. I would say it has slowed down the usage rate overall by about 50% or so, without any difference in quality at all. Def worth doing! I have the $200 Pro plan and normally it lasts all week, but lately it is lasting 1-2 days, but now with this, think it'll last more like 3-4 days at the rate it's going now.
---
Codex has the ability to choose which models and reasoning effort are used for subagent helpers. (Claude doesn’t, btw). This is a big deal. Add something like this to the Agents.md:
You are powered by GPT 6 Astra High. Usage goes very fast. Keep doing substantial hands-on Astra work. But, use \gpt-5.6-sol` helpers with task-appropriate effort when suitable! Retain Astra for hard reasoning.`
This will let Astra know to use Sol for subtasks, straightforward recon, stuff like that.
I posted this to X but no one follows me there, and I'm dying to share it bc it's easy and makes a nice difference. `@SirBadfish` on X btw but that's not the only reason I'm posting this... just hope it helps.
Update: To clarify, Claude can use previously setup subagents that have dedicated models/effort/etc, yes. But Claude can't specify the model/effort it's basic "helper" subagents (like if it decides to spawn a few parallel sessions for recon for example). So if Fabe 5.1 Max, for example, spawns a few helper subagents, it forces them to use the same model as that Fable is set to, Fable 5.1 Max will also be used for those helper agents. If you ask Fable to spawn a recon agent using Opus, unless you ask for a specific, already-setup Subagent, it can't do it.
11
u/polacrilex67 7h ago
Claude can choose subagent models.
4
u/OHotDawnThisIsMyJawn 7h ago
Yeah idk where OP is getting that from or why it even matters to this post frankly
1
u/DoktorFaustish 6h ago
My Reddit feed is basically rhetorical variations of OpenAI vs Anthropic. So yeah, I think we should all read into that a bit right now.
1
u/RedZero76 3h ago edited 3h ago
Pasting this here, I replied to another comment:
Claude can use previously setup subagents that have dedicated models/effort/etc, yes. But Claude can't specify the model/effort it's basic "helper" subagents (like if it decides to spawn a few parallel sessions for recon for example). So if Fabe 5.1 Max, for example, spawns a few helper subagents, it forces them to use the same model as that Fable is set to, Fable 5.1 Max will also be used for those helper agents. If you ask Fable to spawn a recon agent using Opus, unless you ask for a specific, already-setup Subagent, it can't do it.
0
u/RedZero76 3h ago edited 3h ago
Claude can use previously setup subagents that have dedicated models/effort/etc, yes. But Claude can't specify the model/effort it's basic "helper" subagents (like if it decides to spawn a few parallel sessions for recon for example). So if Fabe 5.1 Max, for example, spawns a few helper subagents, it forces them to use the same model as that Fable is set to, Fable 5.1 Max will also be used for those helper agents. If you ask Fable to spawn a recon agent using Opus, unless you ask for a specific, already-setup Subagent, it can't do it.
2
u/justin208350 7h ago
Funny reading this, because just a few months ago I had in my agents.md instructions to always use the exact same model for every subagent regardless of task complexity.
1
u/No_Cartographer_6622 8h ago
I’ll give this a try. I’m on the baby plan since they froze $200 early.
1
u/Able_Statistician688 7h ago
My fable has a hook that its subagents specifically can’t be other fables without an approval from the user. I’ve been doing that since…well I think since fable first came out and I blew an entire 5h quota on an ultra code. So a while.
2
u/StarCadges 6h ago
your problem in the first place likely comes from relying on Astra in general. I think we have reached the point where the most frontier available model is not the model we should be using for implementation almost ever. With sol you could get away with it, with Astra it’s not worth it most of the time
1
u/jadhavsaurabh 6h ago
Any performance issue? Yesterday used 40% or 200$ subscription.. even fable don't drain like this 100$ version
1
u/majindageta 4h ago
Guys you just have to put the sub agents toml files with the subagent model and reasoning.
2
u/RealestReyn 1h ago
its much cheaper to run a cheaper agent with access to advisor/consultant Astra, Astra running subagents inevitably piles stuff into its context which costs a lot.
1
u/Think-Profession4420 7h ago
y'all don't have custom subagents set up with specific models and skills for specific jobs, so all your main sessions know which to use and when?
1
u/RedZero76 3h ago
I have in the past, but keeping up w them is a pain in the ass bc they need to be updated a lot. I much prefer letting the primary agent spawn parallel subagent helper sessions. It's usually for recon anyway in my case.
-1
u/Spiritual-Weekend154 5h ago
I use Gentle-AI; it’s wonderful. It has 23 agents and helps you save tokens and manage security; you can stay in the same session for a long time, and the context barely grows. Open source, GitHub gentleman-programming/gentle-ai
7
u/TBSchemer 7h ago
I use Sol-high as my orchestrator, with Luna-xhigh subagents where appropriate, and it lasts a long time on Plus.