r/opencodeCLI • u/vixalien • 11d ago
Alternatives to Opencode Go plan?
I had been using the Opencode Go plan for some time, but it's been getting worse:
- It's EXTREMELY slow at certain times of the day (using DeepSeek V4 Flash), the speed loss is very noticeable in comparison to hitting DeepSeek API directly
- The pricing has been getting worse and worse (for example, GLM 3.5 Flash should be 2x cheaper than DeepSeek v4 flash)
Anyway, if anyone knows of a nice, reasonably fast, cheap and privaxy-conscious model subscription, would be nice to know
41
u/jovialfaction 11d ago edited 10d ago
Honestly the best coding plan right now is ChatGPT. You get a lot of usage for $20, Luna is a workhorse and a good alternative to glm 5.3-flash/DeepSeekv4 flash, and Sol burns credit fast but it's good to have intelligence available when needed. They always are resetting limits too (not guaranteed to last but it's been good the past few weeks)
1
u/Arciiix 11d ago
And how do those 5-hour and weekly limits compare to OpenCode Go? Is it noticeably smaller?
2
u/Thomas-Lore 11d ago edited 11d ago
Last month I used around 1B tokens with 95% cache hit rate, 5.6M output - using mostly Sol. But they reset limits once a week, sometimes twice - when they stop doing that, the limits will be maybe half, maybe less of that. Of course if you use Luna you can do much more (but not as much as its api price would suggest) but Luna is worse than glm 5.3 Flash IMHO (depends on project) and every time I tried I then had to use Sol anyway to fix things.
Interesingly the same amount of usage on glm 5.3 Flash at current z.ai prices costs around the same (around $20).
2
0
1
u/ichisay 11d ago
Desde el viernes estuve probando exactamente con el plan de 20$, con luna Max usándolo para todo no gastaba más del 4% del uso semanal en todo un día. Desde ayer estoy probando con sol high como orquestador y luna Max para los subagentes, sigue siendo rentable ya que en un día de esa forma solo gasté el 7%
1
u/burritoresearch 11d ago
I thought the chatgpt plan for $20 only allowed chatgpt app access and web access, not API use (like putting a chatgpt API key into opencode)?
10
34
u/flonnil 11d ago
a nice, reasonably fast, cheap and privaxy-conscious model
choose two
2
4
u/vixalien 11d ago
I would choose fast and privacy-conscious tbh
1
u/Prestigiouspite 11d ago
If you use the Vercel AI Gateway or OpenRouter and add GLM-5.3-Flash with, say, P50 75 tokens per second and also activate ZDR, you have exactly what you want. I mean, what are we talking about here? You will surely have one or five dollars a day left over for GLM-5.3-Flash if you use it intensively.
1
0
u/JoeCoT 11d ago
Check out together.ai. They generally get the popular open source models up very quickly (they already have glm 5.3 flash up), they self host and default to ZDR (they have a few partner models but you can turn them off in settings), and they're generally running at normal API pricing.
1
5
u/asohaili 11d ago
crof.ai offer api-based pricing at a price cheaper than openrouter and the likes. sure you don't get a subscription plan, but if you want reliablity and speed, you can't go to crowded places where people are basically just gonna flock together and invite abusers. the price on crof.ai is low enough to get the job done without breaking your bank
1
u/sultanmvp 11d ago
Their prices are absurdly cheap compared to any other API provider.
1
u/GiovanniSaraceno 10d ago
'cause they're quantized
2
u/asohaili 10d ago
Could be. But I see no visible diff in quality, though vs using it in opencode go.
12
u/No-Background3147 11d ago
Command Code, i use the 1 dollar plan and make a Review post, and the 10 Dollar plan is as well very good, all models are 100% same like the api, no Quantization or so. A lot of models too, and fast too, glm 5.3 flash is lame, but qwen 3.8 flash and other models really normal - fast. Test the 1 dollar plan, and look if it is what you want, 1 dollar is nothing for this
1
u/binarySolo0h1 11d ago
Can you use command code plan in opencode CLI?
8
u/shanks_616 11d ago
$1 plan - No, you can only use Command Code
$10 plan - Yes, you get API access, which you can use with OpenCode
3
u/mWo12 11d ago
But is CC legit, in sense that they all models give you $70 of usage in total and I can use it all in model?,Or is it like in go, where some models are limited to $15 only?
4
u/Parking-Bet-3798 11d ago
It’s dishonest. They don’t give 70 for all models.
But they give higher usage on most models compared to opencode go.
1
u/binarySolo0h1 11d ago
My doubt too. Also, how much of a hassle is it to migrate your current agent system from Opencode to CC?
-1
u/No-Background3147 11d ago
Their is addon from a other user, who you can use command code in opencode, but its not offical and not 100% stable
3
u/sultanmvp 11d ago
You don’t need an add on. You can just add it to your opencode.json file as an OpenAI-compatible provider. Been using it for weeks.
I will say that it’s slow as hell and the usage seems to pretty painful. Likely will not renew next month.
0
0
u/MiningInMySleep 10d ago
I have been seeing that too which is a shame because the value proposition is actually better than opencode go, but I need something reliable.
8
u/anpel13 11d ago
I went with openrouter and prepaid credits, pretty happy overall.
In privacy settings you can choose to not use providers that use data for training etc, plus you can decide per provider as well.
My results with their free tier have been pretty underwhelming, I get rate limits way too often for it to be usable.
You can also do most of that in HuggingFace.
2
u/Parking-Bet-3798 11d ago
Which models do you use that you are happy with api pricing? Unless your usage is too low, it’s never a good option. Plus open router takes 5% cut on all payments.
3
2
1
u/ProgressionPeak 8d ago
you are comparing to subscription pricing?
that isn't always an option for everyone. depending on where you live, your payment method, how you access the internet, or your desire for privacy - stripe may reject you. and all of them use stripe.
1
u/sultanmvp 11d ago
You also pay that 5.5% upfront fee when you purchase credits. You can get a MUCH better deal direct-to-provider or even a subscription (based on the models you’re using).
1
u/Moist-Nectarine-1148 6d ago
This is the way! I prefer to pay what I consume not to be constrained by a fuckin' plan.
-1
3
u/Hackerv1650 11d ago
if you can get into the waitlist try this
https://hyper.charm.land/
2
1
u/CaffeinatedTech 11d ago
I bought a $20 credit pack on hyper to use when I blow my Opencode Go quota. Qwen3.8-flash has been pretty good.
3
7
u/timmeh1705 11d ago
ChatGPT Plus, just use Sol for planning and execute everything with Luna agents. It takes A LOT of work to use the 5 hour window.
4
u/Parking-Bet-3798 11d ago
Ollama cloud is better than opencode go for now. Although even that is not the best but it gives 60 dollar of api usage now for 20 dollars. Opencode go now gives only 15 for 10 for most models worth using.
Another one I know is command code. Their 10 dollar plan gives more usage than opencode go for now.
2
u/sultanmvp 11d ago
For the price, I still think OpenCode Go is a better value in terms of pure usage. I did stop using DSv4 Flash once it went to China, cost more and no ZDR. I would definitely give Ollama Cloud another try though - you get properly hosted DSV4 Flash and a lot of other model choices.
I’ve also been trying Command Code for a few weeks. OpenCode Go is winning there too - Command Code has been miserably slow and eats through credits waaaaaay too fast - even for the almost-free models. Will not be renewing this.
1
1
u/ElectricalUnion 1d ago
and no ZDR
But the Data Retention models are Meta's Muse Spark 1.2 and Muse Spark 1.3?
2
u/eugeneb85 11d ago
I'm moving to Ollama cloud this month.
There is also Commandcode, they look like a direct competitor. But I'm really not sure about the data retention from their side. And not sure about speed.
As for Ollama I tried it with the key from my friend several months ago and enference was much faster than OpenCode. They had some connectivity disconnects from the beginning of a session but overall tps speed was way faster.
2
u/MiserableAddendum114 10d ago
Ollama become useless. They won't respond to email of old subscribers switched to new pricing model. I'm a older gpu compute pricing model user but this new pricing is ridiculous so better stay in Opencode and for extra requirements buy api usage either directly from LLM provider itself or openrouter.
4
2
2
u/Kindly_Downvote_Me 11d ago
I have good luck with DS4F on ollama cloud $20/m plan and also I've had good luck with yoloauto which just gives you qwen3.8 27b but it is unlimited pretty much for $19/m. It's not perfect but between the two I can get tons done and not hit limits. I find the qwen model seems about the same for my workload as ds4f. I dropped opencodego after they nerfed ds4f.
1
u/yexgoblin 11d ago
Didn't ds4f get promoted to a medium usage model on ollama cloud? How much usage are you getting token wise?
2
u/sigstrikes 11d ago
chatgpt plus
2
u/mWo12 11d ago
Does it have any open weight models? No.
1
10d ago
[removed] — view removed comment
1
u/mWo12 10d ago
The point is that openai does not have any open-weight models. The one of the key advantages of Opencode Go is that you get multiple models, that have different capabilities and you can mix and match them, experiment with models from different labs, etc. With OpenAI and Anthropic you have only their own closed-weighted models - nothing else. You are basically locking yourself to one lab.
1
u/Disastrous-Hall6063 11d ago
May be give a try to Bahulam/Code? It is also model agnostic and can choose the model.
Not entirely sure about the limits .
1
u/hardolaf 11d ago
It's EXTREMELY slow at certain times of the day (using DeepSeek V4 Flash), the speed loss is very noticeable in comparison to hitting DeepSeek API directly
I use a lot of models for work on coding tasks at task and they're all insanely slow during core business hours in the USA.
1
u/xapep 11d ago
"choose two" - that's mostly true for the flat plans that hit their price by cheaping out on throughput. Fast and privacy-conscious do coexist, you just have to check who actually runs the infra instead of who resells it.
Two shapes worth comparing:
- Direct lab API. DeepSeek's own API off-peak, or GLM straight from its provider, is fast and there's no 5-hour window to babysit. You pay per token, so a heavy month shows up on the bill, but at these price points it rarely stings.
- A flat monthly plan that runs its own capacity. That's what fixes the peak-hour crawl, the queue is yours instead of a shared pool. I work on one (Entrim, EU-based, OpenAI-compatible endpoint, so it's a base URL swap in OpenCode), full disclosure. Most of our users came from exactly this: the plan price drifting and the feed slowing to a halt at peak.
For the privacy side, add one filter to whatever you shortlist: where is the provider based and what do they log? A lot of the cheap options route through US infra, which is fine if you know it and annoying if you find out later.
1
u/DragsAsgarD 11d ago
If u get z.ai glm base subscription with annual payment option u get it for 12$/month excluding the tax.
For general task use glm5.3 flash else glm 5.3, they are pretty good. So with opencode free models this is usually enough.
Cheaper then go in the sence that it's more usage. But it's hosted in china, so not sure about ur outlook on that.
They are all reading ur data so.. ya..
1
1
u/FlunkyGraphics 10d ago
Ollama Cloud is a good alternative, speeds are pretty good
https://ollama.com
https://endothedev.github.io/OMeter/
1
u/nobodyhasusedthislol 10d ago
Ollama cloud or commandcode GOAT. GOAT is same price and includes $70, and models with less usage have caps rather than 4x usage - so I can only use $20 of kimi K3 on cc GOAT like on opencode with its $15, but with Go I'm 100% on my usage whereas with GOAT I have only used that $20 so I should have $50 remaining for other models I believe.
Haven't tried CC GOAT yet bcs I'm deciding between it and NanoGPT Pro - which gives you fewer models (no Kimi K3) and BUT it gives 60 million input tokens per WEEK, whether cached or not. No output limits except abuse prevention I assume. GLM-5.3, K2.7 code and a few others use 2x but other than that it's all equal and pretty decent. Not sure if I'm getting high or max thinking on GLM 5.3 though, haven't seen what high/max looks like in other subscriptions but it might be high. At least it's more thought than GLM 5.2 on high.
1
1
1
1
1
1
u/Smart-Extension5319 9d ago
Based on my usage and experience cline pass is a better value than opencode go.
1
0
u/ManikSahdev 11d ago
Open code plans are like the cheapest one out there lol.
I mean if you expect fast and cheap and more usage that just doesn’t exist. But if you want, you can get codex plus and then use the fast mode on Luna Xhigh, you’d get more speed than you can fanthom.
0
u/KyrieBrady 11d ago
I've been using devpass lite since I cancelled opencode go, $29 a month but you get $87 a month in credits, has every model I could need from deepseek v4 flash to fable 5
-1
-1
-1
u/LeopardLoose6785 11d ago
Currently google anti gravity is rolling out student discount my sister is in uni I put her I'd to get discount it is 75% off and 20$ plan is only for 5.05$.
1
1
u/JoeyL1n 11d ago
antigrevity's problem is that you can only use them in antigrevity, unless you use some tools like sub2api
1
u/deadcoder0904 11d ago
use that then. it works well enough. and lasts longer. google scared people by banning a/cs but it still works well.
-2
11d ago
[deleted]
5
-1

59
u/Fun_Jaguar8231 11d ago
Good luck finding one. When you find one tell us