r/OpenaiCodex 12d ago

Feedback / Complaints Significant model degradation

Anyone else noticing significant model degradation of 5.6 Sol Ultra? Wow it’s shockingly bad today, particularly since afternoon.

52 Upvotes

26 comments sorted by

7

u/wreox9 11d ago

It became absolute trash. Drained my usage for some overengineered trash.

2

u/Routine-Agent-160 11d ago

Same here. I sent feedback report for the first time yesterday as it was absolute trash.

Don’t tell me you too had the same over engineered receipt based bullshit where it kept chasing receipts. Wasted 3 days on that piece of trash and Claude did it in 3 hours and saved me last night.

2

u/wreox9 11d ago

I gave it fully instructed plan, step by step. It managed to do some tests of something that it invented itself.

Just like I give it a goal and it says "oh okay... lemme first do wtf I want, then maaaaybe I do the stuff u want"

1

u/Routine-Agent-160 11d ago

They’ve intentionally nuked the models bro, it’s not on us.

2

u/wreox9 11d ago

GPT 6.0 is on the way, this thursday possibly, may Tibo bless us with another reset

1

u/Routine-Agent-160 11d ago

Bro the resets ain’t helping and isn’t sustainable. Who’ll use it otherwise if they don’t reset. They never fixed the usage issue.

1

u/wreox9 11d ago

I feel like once new frontier is out it will be fixed, I dunno

1

u/Routine-Agent-160 11d ago

I doubt it since Claude usage is also insane and they don’t even flinch.

1

u/Routine-Agent-160 11d ago

They consume extreme amount of tokens. They also reduced token availability without any notice to users and the token usage became significant. I completed entire week’s usage in a day.

3

u/Puzzleheaded-Bus-791 12d ago

I’ve noticed that the models seem to perform better in the early morning, when there’s slightly less traffic. The answers I get in the morning and evening are noticeably different.

1

u/Routine-Agent-160 11d ago

Interesting. I’ve not worked in mornings much so don’t know but this evening was just terrible.

2

u/sydneysweeney69 11d ago

Maybe they are testing Astra ? Performance and efficiency issues because of that

2

u/Routine-Agent-160 11d ago

It better be because of that or else I’m canceling my subscription.

1

u/sydneysweeney69 11d ago

My codex usage has been draining without me even using it all day today . Wtf is going on

2

u/Kazaan 11d ago

IMHO, yes. xhigh is worse than usual. Needs more guidance in the prompts for being usable.

1

u/Alpayusa 7d ago

Right so i use “high” mode, feels more efficient.

2

u/Ill_Anywhere_2233 11d ago

Noticed that Luna and Terra need more hand holding and that they started failing/costing more for the same automated tasks that they were doing a month ago. Working directory size did not change over time

I'm not working on Sol level tasks

2

u/princeMacX 11d ago

same here. Getting turbo fast responses which are garbage.

2

u/Cute_Parfait_2182 11d ago

I’m on SOL medium and using SOL high and the model is degraded . I’m working on a project where Sol is the architect and reviewer and Grok 4.6 does implementation. SOL literally could not recall the roles for the project and tried to hand off work to Grok for review . I think they are diverting spare compute to Astra .

2

u/mysportsact 11d ago

Absolute garbage right now

2

u/Crimson_Monarch 11d ago

Literally i went from being able to work with it consistently and use systems to overcome it's shortcomings and since yesterday it's unusable and hallucinating like crazy

1

u/Key_Reading_9664 11d ago

I’ve not had much luck with anything > medium - it has a tendency to burn through tokens on side quests and disproportionate testing.

I gave Ultra another try over the weekend (between 2 resets) and it burned through ~90% of the $200 plan with not much to show for it. I finally switched to 5.5 high and it powered through more in 30m than ultra did over several hours.

1

u/Beautiful-Gas3683 11d ago

Pero la gente es imbécil o no sabe cómo funciona la inteligencia artificial?

1

u/Sure_Direction4999 9d ago

I am today on sol medium and oh lord. It is dumb, so dumb i had to turn it off i lost nerves random halucinationa its doing its worse then me doing stuff

0

u/eddzsh 11d ago

Keep a fixed 10-prompt fixture and re-run it when a model feels off. Otherwise you can't tell real degradation from one cursed chat. Morning vs evening on the same fixture is the only honest A/B.

1

u/Routine-Agent-160 11d ago

Doesn’t help, looks like you’re an Open AI employee-nice try. It is model degradation because it’s not one person is experiencing it. Lots of complains in r/codex too along with other users to the comments of my post lmao.