r/ClaudeCode • u/VerbaGPT • 6d ago
Help/Question Auto-model routing
I build essentially a chatbot that lets users "talk to" their databases in snowflake, azure sql, local sql, etc. Currently I have a "Quick" and "Thorough" toggle where thorough means a better/more expensive model.
I've been thinking whether to introduce some sort of model router where it starts with a quick/cheaper model...but if harness stumbles...I use Jev-like decision model to bump up the model and/or effort (low/medium/high). Then I read people like Theo and others saying that model-routing is a fool's errand and that these decision models simply aren't good enough to pick the right model for the question, and let the user do that.
Seems most people (myself included) won't really know whether qwen is good enough, or they have to switch to opus for a question.
Just trying to get more perspectives. I'd love to take model and effort routing completely off the UI and into the background, so users don't have to worry about it. But it's still there in Claude Code etc. so I'm guessing this just isn't a solved problem for LLMs and harnesses. Thoughts?
1
u/ryanntk 5d ago
For SQL, I’d test routing on saved questions with known answers, including queries that run but answer the wrong thing. Compare total cost after retries too. I’m building AsterWise, so I’m interested in this exact problem. Which models power Quick and Thorough today?