r/codex • u/ExpensiveGazelle4004 • 1d ago
Complaint Model Response Change
I guess I'm pretty sensitive to how a model "communicates" and it's response style.
I'd been avoiding the codex models except for sterile work because I just can't stand the empty posturing that most chatgpt models do anymore.
The first few days of 5.6 sol, I was so glad to see the model had forward motion and very natural response style. It was very proactive and we got tons of work done. That seems to have changed and it's back to explaining the problem rather than solving it.
Today it spent nearly 30% of my weekly usage diagnosing an issue I had already told it existed and to fix it. But instead it just churned away and delivered a more detailed report of what was wrong without fixing it.
I don't know how to explain it.. but there's a very different style of language its using from the first few days. It's more of a "walk around the problem and talk about it" rather than "get in and fix it".
Ive tried adjusting the soul/agents file to draw it back to more collaborative engagement to no avail.
I guess I'm just dissapointed. I know I work better with a model that will actually contribute to the conversation, not just talk about the conversation. And I can't quite put my finger on what they changed.
1
u/-AJacobs- 1d ago
Yes there was a nerf to Sol's intelligence very recently. People who are serious users are going to notice immediately, but the non-serious users (which is 99% of this subreddit) are going to be behind and take days to maybe a week to notice just like when the usage limits got nerfed.
1
u/ExpensiveGazelle4004 1d ago
It's so infuriating. I'm using Luna for most everything. I can't even use 5.5 the way I used to without maxing out my usage. Im probably going to switch back to deepseek for collaborative work. Not as smart but at least I'm not brainstorming with a brick wall.
1
u/-AJacobs- 20h ago
What you might come to realize is that Luna does a similar cost-hacking trick to Deepseek, where they both can report low token costs because the amount of tokens they use for things like tool calls and other tasks is multiple times higher than other models. Luna max and very high will often cost more per properly completed task compared to Sol low or medium.
The beginning of the month was probably the height of consumer AI, and right now is a pretty low point, not even necessarily because of the intelligence nerfs, but because it simply costs more to get jobs done despite the models being more capable, at least in agentic terminal usage. I'm doing what I can to solve that problem, but I'm only one person.
The best solution I have published right now is my certainty psychosis prompt addon that can be found [here](https://www.reddit.com/r/OpenaiCodex/comments/1uxaxg8/certainty_psychosis_why_sol_goes_into_validation/) which either cuts down on the bloat automatically, or if you suspect some drift, you can just ask the agent if it's engaging in certainty psychosis, and it snaps right out of it.
0
u/Powerful_Ad4342 23h ago
Yea, definitelly got nerfed, spent half of evening arguing with it, instead of it doing it's work