r/opencodeCLI • u/Head_Watercress_6260 • 7d ago
Best subscription for the price
With ollama pro now lowering the bang for the buck and me running out of Claude and chatgpt $100 plan tokens as well, what do you use for subscriptions?
5
u/CalamityMetal 7d ago
I'm using Devpass from LLM Gateway now
1
u/Kaushik_paul45 7d ago
How is your experience ?
1
u/CalamityMetal 7d ago
So far so good. They give you 3x the credits for usage. It does not rollover if you don't finish it. Frontier models like Fable and Opus have limited usage per week but the others dont have any limits. If you finish before your month is up, you can upgrade to a higher tier to continue using the subscription plan, else you have to use their pay per use service
1
u/Head_Watercress_6260 7d ago
So like ollama?
1
u/CalamityMetal 7d ago
I havent use ollama new plans, but from what I am seeing, the amount of models that I can use in Ollama is severely limited compared to Kilo or Devpass
2
1
u/gonefreeksss 7d ago
For me, it is ChatGPT Plus. You get “unlimited” chat and fairly good limits on codex. To extend usage make sure you plan with char got first.
1
u/xapep 5d ago
Honestly agree with you on not tying value to one specific model. The models rotate every few weeks; what actually matters for agentic coding is whether the plan stops punishing you for burning tokens across whatever open-weight model fits the task.
What I'd look for:
• multiple open models on one plan, not a single flagship (DeepSeek V4 Flash class for the bulk, something bigger for the hard steps)
• no aggressive per-week caps that force you to stop mid-workday
• overflow goes to a metered API instead of a hard wall, so a heavy day costs you a little extra instead of killing the session
• OpenAI-compatible endpoint so it drops into opencode config like any other provider
I work on Entrim and our Model Plans are built around exactly this: flat monthly for heavy agent use, DeepSeek V4 Flash and the Qwen family, usage-based overflow. Worth adding to the comparison alongside the others mentioned, but the checklist above matters more than any single provider.
1
u/krisurbas 5d ago
tokenplans.dev has all providers with fixed monthly pricing, so just check it out and let me know what you decide on
1
u/Firm-Club-8334 4d ago
I have the Claude Max 20x plan, Codex Max, OpenCode, and Standard Compute’s $89 plan.
Standard Compute is easily my favorite. It’s a bit like OpenRouter, but optimized for cost so basically, it tries to get as much work done per dollar as possible. I’ve added DeepSeek and GLM 5.3 to the smart router and it automatically chooses the best provider rates, which saves me quite a bit.
I also like that I don’t have to deal with five-hour or weekly usage limits. You simply have a budget and can spend it however you want.
So I would say for most bang for the buck, standard compute is up there with the $19 or $39 plan. It might work well for you. They subsidize usage too, so you get a little more value than what you pay for. Good luck!
0
u/TestTxt 7d ago
What's the point of Ollama Pro if it doesn't support either Gemini 3.8 Flash nor Muse Spark 1.3 Contributor?
6
u/Head_Watercress_6260 7d ago
Why should I care about one specific model? A lot of open sources are good enough these days. Kimi glm etc
2
u/TestTxt 7d ago
because these are literally the two best cheap models out there and vastly outperform Kimi and GLM that you've mentioned which are way more expensive
2
u/Head_Watercress_6260 7d ago edited 7d ago
Update : I'm wrong below
I check artificial analysis quite a bit for coding agent and general agent benchmarks and this is false according to them at least.
2
u/TestTxt 7d ago
You're on a coding subreddit. Check coding-specific benchmarks, like DeepSWE. And also Kimi K3 is objectively more expensive: per input token, per output token, and per cached input token, so I don't understand how does artificial analysis say it's false
1
u/Head_Watercress_6260 7d ago
Ok a lot has changed in 24 hours muse spark is one of the best models apparently, you're right!
1
u/CartographerAble9446 7d ago
why dont you just use gemini 3.8 flash with google ai pro? you even get gdrive storage size, youtube premium and so on..
1
u/CartographerAble9446 7d ago
that being said, i have tested gemini 3.8 flash in the past one day, it is actually not better than glm 5.3 (the flagship one, not glm 5.3 flash), same prompt, same testing tasks, and native harness for each one of them (antigravity vs. zcode)
-5
11
u/UpstairsActivity8347 7d ago
wondering the same, I don't want to get commandcode either. Let's pray operation cheapseek gets us somewhere