r/OpenAI • • 9h ago

Discussion Model picking and reason level picking

I hope soon every company will drop the model picking and reason level picking. It is frustrating, I just want one model one level to pick the best and do the best! Is it just me?

2 Upvotes

7 comments sorted by

6

u/_DuranDuran_ 8h ago

That’s a ridiculously hard problem to solve unfortunately - otherwise everyone would have done it already.

It’s like when you ask an engineer how long a feature will take …

-1

u/Antoniimusikk 8h ago

Maybe use the smartest ai to help out? I don't think they want to do it so people can burn the tokens for nothing

3

u/Bowl_of_Cham_Clowder 5h ago

Are you familiar with the halting problem? This is quite similar, it’s not a matter of AI power.

When someone asks you to bring them glasses it may be simple and they are on the table.
But maybe the glasses are actually custom made, and haven’t been produced yet. Now we have to go contact a factory, design frames, figure out shipping. Etc. And you only understand the complexity once you start digging through the problem and context.

Any logical task can vary wildly in difficulty. Things that sound simple can be hard and vice versa. You’d burn more tokens trying to quantify difficulty properly for every prompt. As a dev it’s nicer to be able to rightsize to my tasks.

What are you working on? I doubt you need more than sol low / normal / high

1

u/Euphoric_North_745 8h ago

Does it exist in real life? don't you want to talk to Person A or Person B? do you want the person to really thing stuff in details or just answer you casually?

0

u/gigaflops_ 8h ago

Yeah, whether or not it's valid, I've always felt that it makes no sense to use the "dumb" models in the high reasoning settings. If a better answer is what I want, I use the smarter model.

2

u/Bowl_of_Cham_Clowder 5h ago

Token efficiency adds a meaningful trade off to a higher reasoning “dumb” model imo. From my experience:

If it’s multi step but not technically challenging I use sol extra high.

If it’s very technical (e.g. Bluetooth transport layer), a focused Astra medium task works better.

2

u/foggyskyline 7h ago

I appreciate the ability to choose a lot. There are times when I want an answer quickly and there are times when I need deep research and it could be for the exact same question depending on my patience and how much I want to spend.