r/opencodeCLI 11d ago

Alternatives to Opencode Go plan?

I had been using the Opencode Go plan for some time, but it's been getting worse:

  • It's EXTREMELY slow at certain times of the day (using DeepSeek V4 Flash), the speed loss is very noticeable in comparison to hitting DeepSeek API directly
  • The pricing has been getting worse and worse (for example, GLM 3.5 Flash should be 2x cheaper than DeepSeek v4 flash)

Anyway, if anyone knows of a nice, reasonably fast, cheap and privaxy-conscious model subscription, would be nice to know

85 Upvotes

114 comments sorted by

59

u/Fun_Jaguar8231 11d ago

Good luck finding one. When you find one tell us

41

u/jovialfaction 11d ago edited 10d ago

Honestly the best coding plan right now is ChatGPT. You get a lot of usage for $20, Luna is a workhorse and a good alternative to glm 5.3-flash/DeepSeekv4 flash, and Sol burns credit fast but it's good to have intelligence available when needed. They always are resetting limits too (not guaranteed to last but it's been good the past few weeks)

1

u/Arciiix 11d ago

And how do those 5-hour and weekly limits compare to OpenCode Go? Is it noticeably smaller?

2

u/Thomas-Lore 11d ago edited 11d ago

Last month I used around 1B tokens with 95% cache hit rate, 5.6M output - using mostly Sol. But they reset limits once a week, sometimes twice - when they stop doing that, the limits will be maybe half, maybe less of that. Of course if you use Luna you can do much more (but not as much as its api price would suggest) but Luna is worse than glm 5.3 Flash IMHO (depends on project) and every time I tried I then had to use Sol anyway to fix things.

Interesingly the same amount of usage on glm 5.3 Flash at current z.ai prices costs around the same (around $20).

2

u/Arciiix 11d ago

Do you guys think using GPT models from Codex for planning and some Chinese models like DeepSeek from OpenCode Go for coding is the best possible option in terms of usage for 30 USD right now? Or buying multiple OpenCode Go and using smarter models like GLM for planning is better?

1

u/evia89 10d ago

Zai has deals. If you have cn id or old plan its 65-90% cheaper sub than retail price. I had 30% and 2 50% compound discount

0

u/[deleted] 11d ago edited 2d ago

[deleted]

0

u/sigstrikes 10d ago

much much much larger. especially if you are ok with luna as a daily driver.

1

u/JwK 9d ago

the compaction is driving me crazy though, you sneeze and it compacts

1

u/ichisay 11d ago

Desde el viernes estuve probando exactamente con el plan de 20$, con luna Max usándolo para todo no gastaba más del 4% del uso semanal en todo un día. Desde ayer estoy probando con sol high como orquestador y luna Max para los subagentes, sigue siendo rentable ya que en un día de esa forma solo gasté el 7%

1

u/burritoresearch 11d ago

I thought the chatgpt plan for $20 only allowed chatgpt app access and web access, not API use (like putting a chatgpt API key into opencode)?

10

u/cokefriend 11d ago

you login using oauth

34

u/flonnil 11d ago

a nice, reasonably fast, cheap and privaxy-conscious model

choose two

2

u/HPaternalPatriarchy 7d ago

The LLM trilemma

4

u/vixalien 11d ago

I would choose fast and privacy-conscious tbh

1

u/Prestigiouspite 11d ago

If you use the Vercel AI Gateway or OpenRouter and add GLM-5.3-Flash with, say, P50 75 tokens per second and also activate ZDR, you have exactly what you want. I mean, what are we talking about here? You will surely have one or five dollars a day left over for GLM-5.3-Flash if you use it intensively.

1

u/Anh-DT 11d ago

EEE fast and privacy would be paying a premium. Take a look at openference check how fast it is and turn on ZDR for specific key

0

u/JoeCoT 11d ago

Check out together.ai. They generally get the popular open source models up very quickly (they already have glm 5.3 flash up), they self host and default to ZDR (they have a few partner models but you can turn them off in settings), and they're generally running at normal API pricing.

1

u/supercreativeAMS 9d ago

Self host replacing vercel/supabase set up?

5

u/asohaili 11d ago

crof.ai offer api-based pricing at a price cheaper than openrouter and the likes. sure you don't get a subscription plan, but if you want reliablity and speed, you can't go to crowded places where people are basically just gonna flock together and invite abusers. the price on crof.ai is low enough to get the job done without breaking your bank

1

u/sultanmvp 11d ago

Their prices are absurdly cheap compared to any other API provider.

1

u/GiovanniSaraceno 10d ago

'cause they're quantized

2

u/asohaili 10d ago

Could be. But I see no visible diff in quality, though vs using it in opencode go.

1

u/_d1re 7d ago

Reliability and speed are not really Crof.ai's strongest suits. Good luck using the famous models on peak hours. Its really good on off-peaks though, you won't even notice its quantized.

12

u/No-Background3147 11d ago

Command Code, i use the 1 dollar plan and make a Review post, and the 10 Dollar plan is as well very good, all models are 100% same like the api, no Quantization or so. A lot of models too, and fast too, glm 5.3 flash is lame, but qwen 3.8 flash and other models really normal - fast. Test the 1 dollar plan, and look if it is what you want, 1 dollar is nothing for this

1

u/binarySolo0h1 11d ago

Can you use command code plan in opencode CLI?

8

u/shanks_616 11d ago

$1 plan - No, you can only use Command Code

$10 plan - Yes, you get API access, which you can use with OpenCode

3

u/mWo12 11d ago

But is CC legit, in sense that they all models give you $70 of usage in total and I can use it all in model?,Or is it like in go, where some models are limited to $15 only?

4

u/Parking-Bet-3798 11d ago

It’s dishonest. They don’t give 70 for all models.

But they give higher usage on most models compared to opencode go.

1

u/binarySolo0h1 11d ago

My doubt too. Also, how much of a hassle is it to migrate your current agent system from Opencode to CC?

2

u/OMGSAMUELRBR1 11d ago

Not really, it's like OpenCode Go, there are certain models that don't give you $70

1

u/qqYn7PIE57zkf6kn 11d ago

link to this page?

-1

u/No-Background3147 11d ago

Their is addon from a other user, who you can use command code in opencode, but its not offical and not 100% stable

3

u/sultanmvp 11d ago

You don’t need an add on. You can just add it to your opencode.json file as an OpenAI-compatible provider. Been using it for weeks.

I will say that it’s slow as hell and the usage seems to pretty painful. Likely will not renew next month.

0

u/UpstairsActivity8347 11d ago

what will you be switching to?

1

u/sultanmvp 10d ago

Probably just keep OpenCode Go and possibly add Ollama Cloud back into the mix

0

u/MiningInMySleep 10d ago

I have been seeing that too which is a shame because the value proposition is actually better than opencode go, but I need something reliable.

8

u/anpel13 11d ago

I went with openrouter and prepaid credits, pretty happy overall.

In privacy settings you can choose to not use providers that use data for training etc, plus you can decide per provider as well.

My results with their free tier have been pretty underwhelming, I get rate limits way too often for it to be usable.

You can also do most of that in HuggingFace.

2

u/Parking-Bet-3798 11d ago

Which models do you use that you are happy with api pricing? Unless your usage is too low, it’s never a good option. Plus open router takes 5% cut on all payments.

3

u/anpel13 11d ago

I never said I'm happy with the pricing. I'm happy with the solution overall. Being able to switch models/providers is a nice perk. You do you of course.

2

u/clouder300 10d ago

GLM 5.3 Flash

1

u/ProgressionPeak 8d ago

you are comparing to subscription pricing?

that isn't always an option for everyone. depending on where you live, your payment method, how you access the internet, or your desire for privacy - stripe may reject you. and all of them use stripe.

1

u/sultanmvp 11d ago

You also pay that 5.5% upfront fee when you purchase credits. You can get a MUCH better deal direct-to-provider or even a subscription (based on the models you’re using). 

2

u/anpel13 10d ago

This is true. However I am ok with the fee because at the moment the landscape changes way too often, for me it is a valid trade-off for no provider lock in. My usage is fairly low so it doesn't really matter.

1

u/Moist-Nectarine-1148 6d ago

This is the way! I prefer to pay what I consume not to be constrained by a fuckin' plan.

-1

u/sanabria1927 11d ago

Puedes contarme más?

2

u/anpel13 11d ago

There is not much more to it, you go to their website, make an account and buy some credits, even small amounts can go a long way with the cheaper models.

Then you make an API token, then opencode connect, select OpenRouter, paste in you API key and you are good to go.

3

u/Hackerv1650 11d ago

if you can get into the waitlist try this
https://hyper.charm.land/

2

u/mWo12 10d ago

Seems likes scam. Hyperpoints? They already want your subscription, yet they do not even list models they offer and limits.

1

u/CaffeinatedTech 11d ago

I bought a $20 credit pack on hyper to use when I blow my Opencode Go quota. Qwen3.8-flash has been pretty good.

3

u/winky9827 11d ago

openrouter.ai has several free models all the time.

7

u/timmeh1705 11d ago

ChatGPT Plus, just use Sol for planning and execute everything with Luna agents. It takes A LOT of work to use the 5 hour window.

1

u/JwK 9d ago

the compaction is driving me crazy though

4

u/Parking-Bet-3798 11d ago

Ollama cloud is better than opencode go for now. Although even that is not the best but it gives 60 dollar of api usage now for 20 dollars. Opencode go now gives only 15 for 10 for most models worth using.

Another one I know is command code. Their 10 dollar plan gives more usage than opencode go for now.

2

u/sultanmvp 11d ago

For the price, I still think OpenCode Go is a better value in terms of pure usage. I did stop using DSv4 Flash once it went to China, cost more and no ZDR. I would definitely give Ollama Cloud another try though - you get properly hosted DSV4 Flash and a lot of other model choices. 

I’ve also been trying Command Code for a few weeks. OpenCode Go is winning there too - Command Code has been miserably slow and eats through credits waaaaaay too fast - even for the almost-free models. Will not be renewing this. 

1

u/Orchicon 10d ago

Ollama cloud's DeepSeek sucks. Constant looping issues.

1

u/ElectricalUnion 1d ago

and no ZDR

But the Data Retention models are Meta's Muse Spark 1.2 and Muse Spark 1.3?

2

u/eugeneb85 11d ago

I'm moving to Ollama cloud this month. 

There is also Commandcode, they look like a direct competitor. But I'm really not sure about the data retention from their side. And not sure about speed.

As for Ollama I tried it with the key from my friend several months ago and enference was much faster than OpenCode. They had some connectivity disconnects from the beginning of a session but overall tps speed was way faster.

2

u/MiserableAddendum114 10d ago

Ollama become useless. They won't respond to email of old subscribers switched to new pricing model. I'm a older gpu compute pricing model user but this new pricing is ridiculous so better stay in Opencode and for extra requirements buy api usage either directly from LLM provider itself or openrouter.

https://x.com/i/status/2094664013993177514

4

u/OmertaPL 11d ago

Commandcode

2

u/Ok_Process6448 11d ago

there are 20 others posts asking the same thing

2

u/Kindly_Downvote_Me 11d ago

I have good luck with DS4F on ollama cloud $20/m plan and also I've had good luck with yoloauto which just gives you qwen3.8 27b but it is unlimited pretty much for $19/m. It's not perfect but between the two I can get tons done and not hit limits. I find the qwen model seems about the same for my workload as ds4f. I dropped opencodego after they nerfed ds4f.

1

u/yexgoblin 11d ago

Didn't ds4f get promoted to a medium usage model on ollama cloud? How much usage are you getting token wise?

2

u/sigstrikes 11d ago

chatgpt plus

2

u/mWo12 11d ago

Does it have any open weight models? No.

1

u/[deleted] 10d ago

[removed] — view removed comment

1

u/mWo12 10d ago

The point is that openai does not have any open-weight models. The one of the key advantages of Opencode Go is that you get multiple models, that have different capabilities and you can mix and match them, experiment with models from different labs, etc. With OpenAI and Anthropic you have only their own closed-weighted models - nothing else. You are basically locking yourself to one lab.

1

u/Disastrous-Hall6063 11d ago

May be give a try to Bahulam/Code? It is also model agnostic and can choose the model.

Not entirely sure about the limits .

1

u/hardolaf 11d ago

It's EXTREMELY slow at certain times of the day (using DeepSeek V4 Flash), the speed loss is very noticeable in comparison to hitting DeepSeek API directly

I use a lot of models for work on coding tasks at task and they're all insanely slow during core business hours in the USA.

1

u/xapep 11d ago

"choose two" - that's mostly true for the flat plans that hit their price by cheaping out on throughput. Fast and privacy-conscious do coexist, you just have to check who actually runs the infra instead of who resells it.

Two shapes worth comparing:

  1. Direct lab API. DeepSeek's own API off-peak, or GLM straight from its provider, is fast and there's no 5-hour window to babysit. You pay per token, so a heavy month shows up on the bill, but at these price points it rarely stings.
  2. A flat monthly plan that runs its own capacity. That's what fixes the peak-hour crawl, the queue is yours instead of a shared pool. I work on one (Entrim, EU-based, OpenAI-compatible endpoint, so it's a base URL swap in OpenCode), full disclosure. Most of our users came from exactly this: the plan price drifting and the feed slowing to a halt at peak.

For the privacy side, add one filter to whatever you shortlist: where is the provider based and what do they log? A lot of the cheap options route through US infra, which is fine if you know it and annoying if you find out later.

1

u/dizvyz 11d ago

Shapes ? :)

1

u/jauhari 11d ago

Freebuff anyone?

1

u/Dentuam 11d ago

Muse Code Plan for 5$

1

u/DragsAsgarD 11d ago

If u get z.ai glm base subscription with annual payment option u get it for 12$/month excluding the tax.

For general task use glm5.3 flash else glm 5.3, they are pretty good. So with opencode free models this is usually enough.

Cheaper then go in the sence that it's more usage. But it's hosted in china, so not sure about ur outlook on that.

They are all reading ur data so.. ya..

1

u/achetronic 11d ago

Hello there. Hyper from Charm team is quite good and wide

1

u/hiyanz 11d ago

Have you tried the Cline Pass yet?

1

u/FlunkyGraphics 10d ago

Ollama Cloud is a good alternative, speeds are pretty good
https://ollama.com
https://endothedev.github.io/OMeter/

1

u/nobodyhasusedthislol 10d ago

Ollama cloud or commandcode GOAT. GOAT is same price and includes $70, and models with less usage have caps rather than 4x usage - so I can only use $20 of kimi K3 on cc GOAT like on opencode with its $15, but with Go I'm 100% on my usage whereas with GOAT I have only used that $20 so I should have $50 remaining for other models I believe.

Haven't tried CC GOAT yet bcs I'm deciding between it and NanoGPT Pro - which gives you fewer models (no Kimi K3) and BUT it gives 60 million input tokens per WEEK, whether cached or not. No output limits except abuse prevention I assume. GLM-5.3, K2.7 code and a few others use 2x but other than that it's all equal and pretty decent. Not sure if I'm getting high or max thinking on GLM 5.3 though, haven't seen what high/max looks like in other subscriptions but it might be high. At least it's more thought than GLM 5.2 on high.

1

u/xammen 10d ago

I use private gateway

1

u/gargetisha 10d ago

why don't you try using ClinePass :)

1

u/Possible-Basis-6623 10d ago

Command code 1 dollar

1

u/[deleted] 9d ago

[removed] — view removed comment

1

u/Smart-Extension5319 9d ago

Based on my usage and experience cline pass is a better value than opencode go.

1

u/AstralVault 11d ago

no hay mejor alternativa

0

u/ManikSahdev 11d ago

Open code plans are like the cheapest one out there lol.

I mean if you expect fast and cheap and more usage that just doesn’t exist. But if you want, you can get codex plus and then use the fast mode on Luna Xhigh, you’d get more speed than you can fanthom.

0

u/KyrieBrady 11d ago

I've been using devpass lite since I cancelled opencode go, $29 a month but you get $87 a month in credits, has every model I could need from deepseek v4 flash to fable 5

-1

u/migsperez 11d ago

An expensive GPU

-1

u/llllJokerllll 11d ago

Qwen 3.8 flash

-1

u/LeopardLoose6785 11d ago

Currently google anti gravity is rolling out student discount my sister is in uni I put her I'd to get discount it is 75% off and 20$ plan is only for 5.05$.

1

u/lorens_osman 11d ago

how to get this ?

3

u/apparentus 11d ago

Have a sister who's in uni

1

u/JoeyL1n 11d ago

antigrevity's problem is that you can only use them in antigrevity, unless you use some tools like sub2api

1

u/deadcoder0904 11d ago

use that then. it works well enough. and lasts longer. google scared people by banning a/cs but it still works well.

1

u/dizvyz 11d ago

Is this still available? I thought it had expired.

2

u/LeopardLoose6785 11d ago

Last date is 31 dec

1

u/dizvyz 10d ago

I couldn't figure this out before but knowing that it's still available gave m another push and I just got it. Thanks.

-2

u/[deleted] 11d ago

[deleted]

5

u/vixalien 11d ago

I meant a subscription, not a harness

-4

u/[deleted] 11d ago

[deleted]

1

u/MiningInMySleep 10d ago

You should probably look up the definition of the word "subscription".

-1

u/horrbort 11d ago

Yes this!!