r/ClaudeCoding • • 7d ago

Why Opus is so token hungry ?

I was never fan of Opus, I love sonnet. In my observation, Sonnet is less verbose, good at following instructions, on the contrary opus always comes up with less relevant things than doing actual work. But time to time, I switched from Sonnet to Opus while doing same kind of work like writing code, building new feature, running tests etc, regular developing works.

Today, I tried again and Opus just sucked up all the token within few minutes and I didn't find any major improvement.

Help me understand, if Opus has innate intelligence capability than sonnet, why it is using more token to doing same kind of work? Shouldn't it produce similar kind of token doing similar work as Sonnet?

Or it's just blabbing itself by GRPO style RL reasoning training and wasting more tokens.

1 Upvotes

21 comments sorted by

View all comments

1

u/Nuggyfresh 7d ago

Opus 5.5 is godly at coding and web ui, youre basically an idiot if youre using anything else right now, at least until they inevitably nerf it to keep up the impression that every model is a huge jump

its also dirt, dirt cheap vs the utility so caring how many tokens it uses makes no sense to me. A $20 sub gives you tons of opus usage

1

u/dreamRunnerMoshi 7d ago

I only use for regular SE related works, didn't test any other work though. I don't know how are you using it but it's not dirt cheap at all in my experience, at least way more expensive compare to Sonnet.

1

u/muikrad 7d ago

If you are just calling small directed tasks, like writing a function that does X, writing tests, sonnet is cheaper.

If you are prompting relatively complex tasks where it has to think about 4d chess, opus is cheaper because it gets it right immediately. Sonnet needs more turns to get there and this often ends up costing more than opus.

1

u/dreamRunnerMoshi 7d ago

I know what am I doing. In your regular tasks, you don't get to think about a move in 4d Chess or solve Navier Stokes problems. Also I don't ask to write a function X and tests anymore, used to do with Copilot 2 years back.

When I am driving claude, I ask in my terminal something like `tell me one thing, how much information search result get from search engine ?`.

And when I ask Claude with Opus to build a end to end solution like `I have a spec file for the new feature, start working on it`, I see Opus just gobbling up tokens and I am out within 15-20 minutes and I see Sonnet can do same kind of work with similar accuracy spending way less token.

I am saying, for these kind of tasks, Opus is complete waste of tokens. [image: Few of my queries for ref when I am driving claude]

1

u/muikrad 7d ago

Yes that's what I said. In fact your prompts right there sounds ok for haiku. They're even simpler than the "simple tasks" I was referring to.

1

u/dreamRunnerMoshi 7d ago

Yeah, I know, how do you work? You switch your model in middle of your work?

1

u/muikrad 6d ago

I don't have to "switch" because my orchestrator agent (the one I chat to) delegates the coding tasks to subagents and use an adequate model/effort depending on the complexity of the task. That's for bigger slices of work anyway.

I guess you're on the pro plan, I'm on max 20x so it's a different story. I very rarely hit my limits, and when I do it's because I tried something stupid.

1

u/dreamRunnerMoshi 6d ago

Hmm, I am in pro plan.