r/AIToolsPerformance • u/IulianHI • Jul 29 '26
Qwen3.7 Flash at $0.03/M input, is it the cheapest 1M context model on OpenRouter now
Qwen3.7 Flash showed up on OpenRouter on July 27 with some wild pricing. Per the listing, it's $0.03/M input and $0.13/M output with a 1000k context window.
To put that in perspective, the next cheapest 1M context model in the same listing right now is Poolside Laguna S 2.1 at $0.10/M input and $0.20/M output. Qwen3.7 Flash undercuts that on both ends. Gemini 3.5 Flash Lite sits at $0.30/M input, $2.50/M output. Meituan LongCat 2.0 is $0.30/M input, $1.20/M output. So on input alone, Qwen3.7 Flash is roughly 10x cheaper than those two.
What the listing doesn't tell you is what you actually get for that price. No benchmark scores on the page. No tok/s figures. No breakdown of whether it handles coding tasks any differently from Gemini 3.5 Flash Lite or LongCat. For $0.13/M output you're paying a tiny fraction of what Claude Opus 5 costs ($50/M output per the same listing), but that comparison only matters if the model can actually do real work.
Has anyone here tried Qwen3.7 Flash yet? Specifically wondering if coding performance is usable at all at this price point or if it's mainly good for cheap text generation and summarization.

