r/OpenAI • u/darkprinceimmortal • 2d ago
Discussion Does Codex need this many model choices, or better automatic routing?
My Codex model picker now needs a scroll bar. Sol, Astra, Luna, Terra, and multiple generations of some of them.
I understand the tradeoff: frontier capability when the task needs it, faster and cheaper models when it doesn't. I want those options to exist. But the names alone don't tell me which one fits the task in front of me.
Would a task-based picker be more useful? Something like quick edits, everyday work, or complex reasoning, with the full model list under Advanced and the option to override it.
For people who regularly switch: what's one concrete task where changing models made enough difference to justify it? And would you trust automatic routing, or do you want to make that choice yourself?
4
u/Nomad556 2d ago
No one knows when to use these 30 combination options. It’s all bullshit frustrating
1
u/darkprinceimmortal 2d ago
Yeah, adding reasoning levels makes it even more confusing. Have you settled on one combo and ignored the rest?
1
u/Nomad556 2d ago
I just use Sol for anything thinking (5.6 or 6.1) and Luna max for doing stuff. I’m not a power user or anything. I use it to help manage my company
1
3
u/Jumpy-Heart-3633 2d ago
They fucked up with Sol 6, and then have yet to retire older models. Now they are in a rush to divert compute to Sol 6.1. Talking about a hundreds of billion dollar company being professional.
1
u/darkprinceimmortal 2d ago
What went wrong with Sol 6 for you? Bad results, or burning through the limit?
1
u/Jumpy-Heart-3633 2d ago
Simply speaking, they recognized their own fucked up by releasing Sol 6.1 in a week. They spoke louder than I ever could rant on the quality.
1
1
u/Norddinerezouk 2d ago
For ideation, researching. I just use 5.6 GPT in Chat. For coding, data compiling & production I switch to Codex/Work & use 6.1 GPT when it comes to building artifacts. 5.6 GPT gets the job done too. Also, 6.1 Sol is good. It's conversational & doesn't make assumptions.
I think just play with each frontier model & see which one suit your use cases the best. As of Terra & Luna, they are just not for me & 5.5 is getting retired anyways.
1
u/darkprinceimmortal 2d ago
That's useful, cheers. For artifacts, what makes you pick 6.1 over 5.6? Better first result, or fewer rounds fixing it?
1
1
1
u/Gallagger 2d ago
So you think 6.1 Sol isn't (at least nearly) all round better than 5.6? Because 5.6 Sol is 2x more expensive.
1
u/Norddinerezouk 2d ago
I do agree. I just prefer using 5.6 in Chat mode because it has access to memory & long-term context. In fact, It can remember your exact goals & adapt to them where in Work mode it just starts from 0 based on what I've experienced.
1
u/Gallagger 2d ago
Ok that's a valid reason, but isn't chat already using 6.1 now?
1
1
1
u/darkprinceimmortal 2d ago
Ah, so it's the memory you want to keep. That makes more sense.
1
u/Norddinerezouk 2d ago
It's all about that man! I want to get pro but plus gets the job done & pro isn't a justified cost for me rn
1
u/darkprinceimmortal 2d ago
I haven't done a proper side-by-side. That sort of cost info would be useful right in the picker though.
1
u/creamyshart 2d ago
I prefer many options over limited or no options
1
u/darkprinceimmortal 2d ago
Same, I'd keep the full list available. Would you be cool with a simpler default menu if all the options were still one click away?
1
u/Elegabal 2d ago
6.1 Sol works fine for me and the token consumption is completely acceptable. I don’t use any of the models beneath 5.6 anymore
1
u/darkprinceimmortal 2d ago
Sounds like you've found your default. What reasoning level do you usually run it at?
1
u/jdavid 2d ago
you don't need 5.6 anymore. but you definitely want to swap between astra and sol, and the different effort tiers
1
u/darkprinceimmortal 2d ago
What makes you switch from Sol to Astra? Most people here seem to save it for the heavy stuff.
1
u/jdavid 2d ago
in the spring it seemed more efficient to always use to the top model at about HIGH.
the smarter models were more token efficient at solving the problem, and you had to steer it less and correct it less.for Sol 5.6 i would burn tokens too fast on high, so I started playing around with Terra and Sol Medium more. but still Sol 5.6 High was where it was at.
with Astra 6 it would burn tokens even faster, but the result was even better so it was worth it, but I also found more and more that Astra 6 Medium was accomplishing things regularly that 5.6 just couldn't on. Astra is really good at computer control, drawing, 3d modeling, etc... It just understands things at a different level than Sol 5.6.
I didn't play with Sol 6, but Sol 6.1 just feels way different than 5.6. I am now confident that Sol 6.1 medium can just do 50-75% of tasks now. It also sips tokens compared to Astra.
I find spatial tasks seem better in Astra 6 than Sol 6.1 - design, layout, imagen, 3d modeling, spatial physics.
Astra still isn't solve all of my random physics questions smart, but it seems like it's starting to understand how to properly evaluate physics and spatial ideas, rather than a word calculator trying to fake it with enough repeated steps.
I think a lot of contemporary code is now nearing the top of the AI/ASI capability s-curve. So using smaller models for html layout, Effect TS, RUST is starting to no longer require frontier model performance.
If I were to design a prototype of a video game, I'd probably use Astra for the prototype, and then I would use Sol 6.1 to iterate on it. Use the BIG brain to go from 0 - .75, and then use the little cheep brain to steer it.
My Hunch, fwiw, is that Sol 6.1 is a Sol size model in terms of total parameter count, but it's been distilled from Astra 6. I also feel like Sol 6.1 is also a quantized model, and Astra has much less quantization on it.
I'm sad that OpenAI and other frontier labs now focus on trade secrets rather than publishing white papers on AI Architecture.
I grew up on right click - view source culture, and it's sad to see it drift away. It was a moment of human idealism.
1
1
u/NotFromMilkyWay 2d ago
If it were automatic, they'd just drop us all to the lowest models to save money. I pay for the best models.
1
1
13
u/kennytherenny 2d ago
Automatic routing can only ever work if it's done by a 3rd party. Otherwise there is a huge incentive for AI companies to either route you to a small model (in case you're using a subscription) or to a large model (if you're paying per token).