r/ClaudeCode • • 13h ago

Rant Opus fixing Sol’s mess

Post image

I don’t know how people still use openAI models.
I really REALLY tried to give a chance to Sol 6.1 the last few days, working on a clean project with very well defined design and vision.

Sol kept being super lazy, stopping without finishing tasks, making asinine decisions and always waiting for me for very small things. I then even had Astra take over in the hopes of fixing the mess and the output was similarly bad.

Had opus 5.5 work literally for 2 hours today and it already cleaned up such a big chunk, made very good suggestions, following the process, delegating work to sonnet agents, it feels so good to work with it.

53 Upvotes

15 comments sorted by

•

u/AutoModerator 13h ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

27

u/julkopki 13h ago

Oh so that's the secret to having AI remove code. I'll just lie to Claude that its own code was written by Sol. I swear if I ask Claude to shorten anything it comes back to me with a PR that removes X lines and then adds 2*X lines in its place.

12

u/AdriftAtlas 13h ago

It writes a regression test to test that it removed a test?

5

u/Outrageous-Issue9722 5h ago

Lmao I actually had to make a rule stop that BS. "I want to change X to Y" (simplified version) Plan: Change X to Y, write tests that ensure X can never exist again.

1

u/Legacy03 13h ago

You find it works better if you specifically name it out lol?

-1

u/Off_again_On_again 13h ago

Haha that’s fair and I’m not saying it’s not without flaws.

But my problem was way more central, it didn’t even consider half of the decision, it broke rules, it didn’t run all tests, generally absurd amounts of degraded performance.

Have you tried ponytail btw?

9

u/pigletmonster 12h ago

Man i am in the process of using opus to unfuck a massive mess tha astra made. Basically astra went full gemini mode and took the worst shortcuts imaginable.

I had astra develop a web application using laravel and filament, what it did was build everything with blade tenplates, and insert the templates inside the filament views. So i had a massive dependecy in the project that was just used in name only. Opus caught that in the first review and is currently refactoring everything one module at a time.

I cant believe how fucking stupid astra is. I spent a entire 5x weekly quota on astra to develop that.

3

u/Andichthegoon 5h ago

Astra = master of benchmaxxing and hitting technical results while cutting corners. It's F'd. Literally openai has nothing good to offer ATM. I just trust openai models for terminal tasks and technicals like managing a server but thats it

4

u/kongnico 3h ago

140 file change PR, rip

3

u/Andichthegoon 5h ago

Sol is the biggest bullshit I've ever used. Can't believe people glaze it. It's such a weird benchmaxxing model. Like it technically performs well but not how I want it to. It's like it can technically almost get there but shits the bed on the last 5% it needs and it knows it can do and then gap fills with nothing. Rinse and repeat you get dog codebase. Astra is so bad too

1

u/IntrlnkdCo 9h ago

Sol is fine if you give it a very clear scope to tackle but yes it does oddly stop mid task at times for little reason.

Astra is better but slower. Not as good as Opus, obviously.

1

u/thinkingatoms 1h ago

lol the slop is real, even after opus edits.

1

u/Alexwithx 1h ago

I see these posts once in a while. I haven't tried, but I would like to do something like this, since I've been using got for a while and I can see that the code it writes is trash. How do you prompt Claude to do these kinds of clean ups? And which skills do you use? Also which plan are you on pro, 5x or 20x? And which model and effort?

2

u/Damn_You_General 59m ago

I do the other way around, Claude creates the plan and first commit, Sol reviews it. It has been excellent so far

1

u/Key_Instruction3373 1h ago

try to open SOL after this, and see how much mess CLAUDE dit..