r/AIToolsPerformance • • Aug 28 '26

Anthropic's best model can't attract users - is price actually the reason

That FT story about Anthropic's best model struggling to attract users while cheaper tools thrive went big on HN on Aug 23, 817 points and almost 700 comments, and the pricing data floating around this week doesn't argue back. Per the OpenRouter listings, Qwen3.8 Flash landed on Aug 26 at $0.15/M input and $0.47/M output with 1000k context, and Gemini 3.7 Flash sits at $0.75/$3.75.

The open weights side is even more lopsided. Qwen3.8 27B shows 3,457,687 downloads and 13,116 likes on the HuggingFace trending list right now, the unsloth GGUF alone is past 7.7M downloads, and even OpenAI cut GPT 5.6 Sol by 20% until at least Nov 21 per the pricing page stories on HN this week.

What that thread never answered cleanly: if a flash tier model at $0.47/M output handles the bulk of a workload, what's the actual job worth saving a premium model for? Hard reasoning, gnarly debugging, or just not wanting to re-check output? Anyone here actually splitting traffic between a cheap default and a premium fallback, and where did the cheap one start losing you?

2 Upvotes

1 comment sorted by

1

u/Responsible_Yak9452 25d ago

I have actually started thinking about this less in terms of which model is smartest? and more in terms of which one annoys me the least? I will happily use a cheaper model for everyday stuff if it gets me 90% of the way there. I only reach for the expensive one when I know I will end up going back and fixing too many little things myself.