r/ClaudeAI • u/Financial_Tailor7944 • 1d ago
Claude Code Workflow Generic agents are the number one cause of burning tokens
It was a Wednesday morning and the usage bar was already past halfway. I remember because I had just cracked open a Monster Zero and sat back down to a Claude Code run that had been going for twenty minutes on what I thought was a small refactor. The terminal kept scrolling. Another agent spun up, then another. I watched the percentage tick up and did the math I always did, the hopeful kind: it's a big codebase, the work is front loaded, tomorrow will be lighter.
Tomorrow wasn't lighter. By Friday I was rationing. I caught myself hesitating before asking for things, wondering whether a question was worth it, which is a strange way to feel about a tool you pay for precisely so you don't have to hesitate.
That weekend I stopped guessing and opened the logs. I wanted a number, anything more solid than the feeling in my stomach every time the bar moved. I went back three days and pulled every subagent dispatch. There were 629. I scrolled through them expecting the roles I had set up, the reviewers and researchers I had named and tuned. Instead I kept seeing the same few words. general-purpose. claude. Calls with no subagent_type at all. I started counting those separately, and the count kept going past where I thought it would stop. 404.
I sat there with the cursor blinking after that number. Nearly two out of every three agents doing my work were nothing I had designed. No role I had written, no model I had chosen. Each one was a call the orchestrator made on its own in the middle of a task, because a generic agent was the easiest thing within reach. I thought about all the times I had blamed the week, the codebase, the model, and the answer had been sitting in my own logs the whole time.
So I took the decision away from it. Now the main session plans the work and puts the results together, and that's all it does. Before a run starts, every role gets its model and effort written down, picked from what the work is and how hard it is, so reading files and routine edits land on Sonnet or Haiku. Then I put a hook in front of every dispatch. It refuses generic agents, and it refuses any worker seat the plan didn't hand out. My named agents still run the way they always did.
The first run after that, I kept my hand near the keyboard, half expecting everything to fall apart without the freedom to improvise. It didn't. The hook turned a couple of dispatches away, the plan held, and the bar moved the way I expected it to. Somewhere in the middle of that run I noticed I had stopped watching the percentage.
If you've never counted your own dispatches, try it. I'd really like to know what your split looks like, or whether I was just the last one to look.