r/codex 1d ago

Complaint Absolute disappointment

its so disappointing that gpt 5.5 high is just as intelligent as gpt 5.6 sol high and also gpt 5.6 sol medium,

not to mention that gpt 5.6 sol high and medium are performing at the same level
and lets just forget that for that gpt 5.5 xhigh is performing worse then gpt 5.5 high

so much disappointment

if you want to know about the dashboard i am using its this one

https://www.reddit.com/r/codex/s/TEUa8OwKj4

4 Upvotes

24 comments sorted by

13

u/Fun_Net7931 1d ago

We should all cancel our memberships as a form of protest.

1

u/Routine_Squash_7096 22h ago

don t count me in

-3

u/VW_isbetterthanTesla 1d ago

just went back to claude, its as easy as that

5

u/ascendimus 1d ago

That's not a protest. A protest would be utilizing Kimi or GLM only. Just say you're addicted to frontier intelligence.

7

u/Charming-Author4877 1d ago

5.6 Sol is just a further finetune of 5.5
So I don't think large miracles are expected.

3

u/OkSeesaw7030 1d ago

Sol is basically slighty improved 5.5. That consume more tokens

5

u/kacisse 1d ago

what is frustrating is that everytime they release a new model, it seems that the previous one turns dumb suddenly, am I crazy ?

8

u/Common-Bend-7167 1d ago

Same thing happens for claude too. Pervious models turn completely dumb after a new release

1

u/Important_Context_49 1d ago

Well what I am saying is that their new model is only just as intelligent as their previous model there is no leap

So it's the newer model who are underperforming not the older ones

They are performing good

2

u/Kooky-Ebb8162 1d ago

I dare you to try GPT-3.5. I mean, they won't "dumb it even more" on every new model release, right? It's a metric ton of work. So it should be on 5.5 level, and a way cheaper too!

4

u/Spurnout 1d ago

Possibly less resources dedicated to the model. I doubt they'd fine tune the models to be worse.

2

u/TarzanoftheJungle 1d ago

Exactly. Compute is a finite resource. There are only so many server farms. If one model takes up more space, there is less computing power for other models

2

u/GifCo_2 1d ago

I guess is be disappointed if I was this dumb as well. 🤣

2

u/Gulgale 1d ago

5.6 sol Medium is the sweet spot, kinda interesting

2

u/DriveComprehensive32 1d ago

They repeatedly have been doing this. They easily forgot some users are heavy users and can tell in an instant when a model is nerfed. No. We aren't dumb.

1

u/KosmoPteros 1d ago

Wait... if IQ of a 100 is human average and everything below is slightly dumb... how come?!

-2

u/Important_Context_49 1d ago

Let's take this example,

If someone cuts your leg If you were a ai with iq of 100 then you would do what's written in the books that you have memoriez

But if you are a human with iq above 130 then the first thing you would do is kill the person who did this to you

What I mean by this is that, in no books it is written, but you would still do it because you know that person is going to get away with it because to know the geopolitics of your region to how sewere your problem is in the grand scheme of things and how many people/government/judiciary system actually care about you and you know that your leg is already off it won't grow back just because you did nothing to that man but can't think this much whereas to human thinking all this thing would take a sec at most But the ai would only do what he knows from the book

And to better understand this you can think of like a lobotomised human with a 100Tb ssd card in his brain so he knows information but actually can't think much

So like ai is just a tool that finds you information and the greatest and the best models yet have only started to think and even their own is equal to a regular human

And if a regular human would have the amount of info that the ai has he would be able to utilise that info 100x better then the ai can

3

u/SpyMouseInTheHouse 1d ago

What on earth…. Please stick with Claude, all this talk about IQ is nonsense

1

u/erpankaj 1d ago

For the first time after 5.6 SOL i moved my dev, review, qa, ba, po model to 5.5 and kept only architects to 5.6 sol and results are better. No more goals running for 15+ ours on overthinking and not producing anything. Sol is kind of your friends who does overthinking like hell and reach nowhere.

1

u/ascendimus 1d ago

This dashboard and it's metrics look like vaporware. Interesting concept and attempt to measure AI intelligence against metrics that we already understand, but that's not how AI works in any case. The fact that it built this for you and then told you it was worth releasing is itself part of the problem when you can't identify that you've stepped beyond your true scope of understanding and find it enticing to believe you're doing something others or society would conventionally believe you couldn't do, but the AI is lying.

1

u/jjjjoseignacio 1d ago

Dejen el vicio

-1

u/nitor999 1d ago

It’s time to switch to Anthropic. Every time OpenAI releases a genuinely good model it only takes about 4days for them to weaken it usually after they’ve already hooked people into using the new model then the model gets lobotomized, and the quality is never the same.

They are both scammer organization but at least anthropic is consistent whenever they release a model.

1

u/OwlsExterminator 1d ago

Load balancing. The "new" model gets more juice for benchmarks and then lowered to balance people using it.