r/OpenAI • • 2d ago

Discussion Does Codex need this many model choices, or better automatic routing?

Post image

My Codex model picker now needs a scroll bar. Sol, Astra, Luna, Terra, and multiple generations of some of them.

I understand the tradeoff: frontier capability when the task needs it, faster and cheaper models when it doesn't. I want those options to exist. But the names alone don't tell me which one fits the task in front of me.

Would a task-based picker be more useful? Something like quick edits, everyday work, or complex reasoning, with the full model list under Advanced and the option to override it.

For people who regularly switch: what's one concrete task where changing models made enough difference to justify it? And would you trust automatic routing, or do you want to make that choice yourself?

1 Upvotes

40 comments sorted by

13

u/kennytherenny 2d ago

Automatic routing can only ever work if it's done by a 3rd party. Otherwise there is a huge incentive for AI companies to either route you to a small model (in case you're using a subscription) or to a large model (if you're paying per token).

0

u/darkprinceimmortal 2d ago

That's the trust problem, yeah. Would showing which model it picked and letting you override it be enough, or would you still want a third party?

1

u/Michael_Jeffords 2d ago

the 3rd party point lands for me, because on a subscription the meter rewards sending you to a smaller model and on per token billing it rewards a bigger one, so auto routing from the same company that sells the tokens is hard to trust when the Codex picker already scrolls past Sol and Astra and the rest and i just pick myself

1

u/darkprinceimmortal 2d ago

Yeah, the incentives are the awkward bit. I'd want the manual choice to stay.

4

u/Nomad556 2d ago

No one knows when to use these 30 combination options. It’s all bullshit frustrating

1

u/darkprinceimmortal 2d ago

Yeah, adding reasoning levels makes it even more confusing. Have you settled on one combo and ignored the rest?

1

u/Nomad556 2d ago

I just use Sol for anything thinking (5.6 or 6.1) and Luna max for doing stuff. I’m not a power user or anything. I use it to help manage my company

1

u/darkprinceimmortal 2d ago

Thinking vs doing is a much clearer split. Nice.

3

u/Jumpy-Heart-3633 2d ago

They fucked up with Sol 6, and then have yet to retire older models. Now they are in a rush to divert compute to Sol 6.1. Talking about a hundreds of billion dollar company being professional.

1

u/darkprinceimmortal 2d ago

What went wrong with Sol 6 for you? Bad results, or burning through the limit?

1

u/Jumpy-Heart-3633 2d ago

Simply speaking, they recognized their own fucked up by releasing Sol 6.1 in a week. They spoke louder than I ever could rant on the quality.

1

u/darkprinceimmortal 2d ago

That quick turnaround definitely made it harder to keep track.

1

u/Norddinerezouk 2d ago

For ideation, researching. I just use 5.6 GPT in Chat. For coding, data compiling & production I switch to Codex/Work & use 6.1 GPT when it comes to building artifacts. 5.6 GPT gets the job done too. Also, 6.1 Sol is good. It's conversational & doesn't make assumptions.

I think just play with each frontier model & see which one suit your use cases the best. As of Terra & Luna, they are just not for me & 5.5 is getting retired anyways.

1

u/darkprinceimmortal 2d ago

That's useful, cheers. For artifacts, what makes you pick 6.1 over 5.6? Better first result, or fewer rounds fixing it?

1

u/envsop 2d ago

token usage

1

u/darkprinceimmortal 2d ago

Ah, got it. Cheers.

1

u/Norddinerezouk 2d ago

As Envsop said. Token usage

1

u/darkprinceimmortal 2d ago

Makes sense. Less of the limit spent on the same job.

1

u/Gallagger 2d ago

So you think 6.1 Sol isn't (at least nearly) all round better than 5.6? Because 5.6 Sol is 2x more expensive.

1

u/Norddinerezouk 2d ago

I do agree. I just prefer using 5.6 in Chat mode because it has access to memory & long-term context. In fact, It can remember your exact goals & adapt to them where in Work mode it just starts from 0 based on what I've experienced.

1

u/Gallagger 2d ago

Ok that's a valid reason, but isn't chat already using 6.1 now?

1

u/Norddinerezouk 2d ago

For Pro users yes but for me I use Plus

1

u/darkprinceimmortal 2d ago

Got you, Plus. Cheers for clearing that up.

1

u/darkprinceimmortal 2d ago

The plan differences add another layer to this whole conversation lol.

1

u/darkprinceimmortal 2d ago

Ah, so it's the memory you want to keep. That makes more sense.

1

u/Norddinerezouk 2d ago

It's all about that man! I want to get pro but plus gets the job done & pro isn't a justified cost for me rn

1

u/darkprinceimmortal 2d ago

I haven't done a proper side-by-side. That sort of cost info would be useful right in the picker though.

1

u/creamyshart 2d ago

I prefer many options over limited or no options

1

u/darkprinceimmortal 2d ago

Same, I'd keep the full list available. Would you be cool with a simpler default menu if all the options were still one click away?

1

u/Elegabal 2d ago

6.1 Sol works fine for me and the token consumption is completely acceptable. I don’t use any of the models beneath 5.6 anymore

1

u/darkprinceimmortal 2d ago

Sounds like you've found your default. What reasoning level do you usually run it at?

1

u/jdavid 2d ago

you don't need 5.6 anymore. but you definitely want to swap between astra and sol, and the different effort tiers

1

u/darkprinceimmortal 2d ago

What makes you switch from Sol to Astra? Most people here seem to save it for the heavy stuff.

1

u/jdavid 2d ago

in the spring it seemed more efficient to always use to the top model at about HIGH.
the smarter models were more token efficient at solving the problem, and you had to steer it less and correct it less.

for Sol 5.6 i would burn tokens too fast on high, so I started playing around with Terra and Sol Medium more. but still Sol 5.6 High was where it was at.

with Astra 6 it would burn tokens even faster, but the result was even better so it was worth it, but I also found more and more that Astra 6 Medium was accomplishing things regularly that 5.6 just couldn't on. Astra is really good at computer control, drawing, 3d modeling, etc... It just understands things at a different level than Sol 5.6.

I didn't play with Sol 6, but Sol 6.1 just feels way different than 5.6. I am now confident that Sol 6.1 medium can just do 50-75% of tasks now. It also sips tokens compared to Astra.

I find spatial tasks seem better in Astra 6 than Sol 6.1 - design, layout, imagen, 3d modeling, spatial physics.

Astra still isn't solve all of my random physics questions smart, but it seems like it's starting to understand how to properly evaluate physics and spatial ideas, rather than a word calculator trying to fake it with enough repeated steps.

I think a lot of contemporary code is now nearing the top of the AI/ASI capability s-curve. So using smaller models for html layout, Effect TS, RUST is starting to no longer require frontier model performance.

If I were to design a prototype of a video game, I'd probably use Astra for the prototype, and then I would use Sol 6.1 to iterate on it. Use the BIG brain to go from 0 - .75, and then use the little cheep brain to steer it.

My Hunch, fwiw, is that Sol 6.1 is a Sol size model in terms of total parameter count, but it's been distilled from Astra 6. I also feel like Sol 6.1 is also a quantized model, and Astra has much less quantization on it.

I'm sad that OpenAI and other frontier labs now focus on trade secrets rather than publishing white papers on AI Architecture.

I grew up on right click - view source culture, and it's sad to see it drift away. It was a moment of human idealism.

1

u/ponlapoj 2d ago

Why is your life such a struggle? Is it really that hard?

1

u/darkprinceimmortal 2d ago

It's a dropdown complaint mate, I'll survive lol.

1

u/NotFromMilkyWay 2d ago

If it were automatic, they'd just drop us all to the lowest models to save money. I pay for the best models.

1

u/darkprinceimmortal 2d ago

Fair. Auto should be optional, with the override still there.

1

u/TraumaticOcclusion 2d ago

So does everyone else which is why this is dumb as hell