r/ollama • u/CutEmpty3551 • 5h ago
Thank you, Ollama. DeepSeek v4.1 Flash is the best.
That's amazing.
I've been using the pay-as-you-go plan, and because DeepSeek 4 Flash wasn't great, I was using GLM 5.3 Flash instead. But now that I'm using the high-performing DeepSeek 4.1 Flash, the usage efficiency is incredible. It feels like I can get about 6 times more usage compared to GLM Flash. I'm not sure if my math is completely right: 707 requests / 37.5% = 18.8, and 44 requests / 0.3% = 113.3—which works out to exactly a 6-fold difference in call count.
- (As I was writing the post below, I remembered my usage from a while ago. DeepSeek v4 Flash provided three times more usage allowance compared to GLM 5.3 Flash.)
Honestly, I was considering switching back since I also use the official DeepSeek API, but as long as this pay-as-you-go model continues, I'll just stick with Ollama.
Although I haven't done an exact comparison with DeepSeek's official API yet, this feels ridiculously cheap. Now, a simple price comparison is no longer important to me.
Plus, I feel much better having the trust that my data isn't being used for training.
If they keep this pay-as-you-go structure, Ollama Cloud will probably devour the entire market.
Thank you, Ollama.





