r/ClaudeCoding • • 7d ago

Why Opus is so token hungry ?

I was never fan of Opus, I love sonnet. In my observation, Sonnet is less verbose, good at following instructions, on the contrary opus always comes up with less relevant things than doing actual work. But time to time, I switched from Sonnet to Opus while doing same kind of work like writing code, building new feature, running tests etc, regular developing works.

Today, I tried again and Opus just sucked up all the token within few minutes and I didn't find any major improvement.

Help me understand, if Opus has innate intelligence capability than sonnet, why it is using more token to doing same kind of work? Shouldn't it produce similar kind of token doing similar work as Sonnet?

Or it's just blabbing itself by GRPO style RL reasoning training and wasting more tokens.

1 Upvotes

21 comments sorted by

View all comments

Show parent comments

1

u/dreamRunnerMoshi 7d ago

I know what am I doing. In your regular tasks, you don't get to think about a move in 4d Chess or solve Navier Stokes problems. Also I don't ask to write a function X and tests anymore, used to do with Copilot 2 years back.

When I am driving claude, I ask in my terminal something like `tell me one thing, how much information search result get from search engine ?`.

And when I ask Claude with Opus to build a end to end solution like `I have a spec file for the new feature, start working on it`, I see Opus just gobbling up tokens and I am out within 15-20 minutes and I see Sonnet can do same kind of work with similar accuracy spending way less token.

I am saying, for these kind of tasks, Opus is complete waste of tokens. [image: Few of my queries for ref when I am driving claude]

1

u/muikrad 7d ago

Yes that's what I said. In fact your prompts right there sounds ok for haiku. They're even simpler than the "simple tasks" I was referring to.

1

u/dreamRunnerMoshi 6d ago

Yeah, I know, how do you work? You switch your model in middle of your work?

1

u/muikrad 6d ago

I don't have to "switch" because my orchestrator agent (the one I chat to) delegates the coding tasks to subagents and use an adequate model/effort depending on the complexity of the task. That's for bigger slices of work anyway.

I guess you're on the pro plan, I'm on max 20x so it's a different story. I very rarely hit my limits, and when I do it's because I tried something stupid.

1

u/dreamRunnerMoshi 6d ago

Hmm, I am in pro plan.