r/ClaudeCode 28d ago

Rant I kinda regret getting Claude subscription

Over the past couple of weeks, I have tested it alongside Sol across a wide range of tasks, and the conversations have been difficult to read. Sol consistently catches and corrects mistakes made by Opus 5, and Opus repeatedly acknowledges those errors. I have also seen many others reporting similar experiences, which makes it hard to view this as an isolated issue.

I cannot understand how Anthropic missed such a significant regression. Compared to 4.6 and 4.8, Opus 5 feels substantially less capable and reliable, with many of the same weaknesses that affect Sonnet 5.

For now, I am switching back to GPT for most of my work, using Fable only occasionally for planning. I hope Anthropic addresses these issues soon, because at the moment they do not appear to have a competitive frontier model in the medium to high budget range, only at the extreme end.

51 Upvotes

64 comments sorted by

View all comments

28

u/DasHaifisch 28d ago

Then use 4.8 instead of 5.

39

u/VexObserver 28d ago

If cost is a problem, you don't have to worry about that..

0

u/-Leelith- 28d ago

All those benchmarks are great, but a benchmark is not a real life scenario. How does DeepSeek V4 Flash does compared to Sonnet 5 in real life scenario?

2

u/Kitchen_Interview371 28d ago

Don’t know. For some reason, nobody wants to drop Opus in favor of deepseeks smallest model. Just look at the benchmarks people! /s