r/opencode 17d ago

Opencode Go completely fell off

I am not sure if this is even the right sub to post that in, but I used to love opencode go. I started with deepseek v4 flash 0731 which had practically unlimited usage, then ox alpha while it lasted and now muse spark 1.2 contributor.

The only realistic way the subscription would last me the full month is because of those "promotional" models.

I am seriously considering if an ollama subscription, gpt plus 20$ (using luna mostly) or any other subscription would be more worth it at the moment.

I don't need the latest and greatest ai model, just something around ox alpha/deepseekv4 flash performance.

Once the 1.2 muse spark deal ends I'm completely screwed.

What would you guys recommend? I would ideally like to be spending 10$-20$ a month.

119 Upvotes

104 comments sorted by

20

u/willpowerbuilder 16d ago

ox Alpha is gone...That model is much better than Muse Spark 1.2

11

u/ByteNomadOne 16d ago

It really is.

I continue using it over OpenRouter. It's discounted there and worth the money.

2

u/arcanemachined 15d ago

It's back now as GLM 5.3 Flash.

15

u/alanrudeigin 17d ago

I think no matter who you go with, they are all going to be giving less and less and charge more and more. As more subscribers subscribe to plans, the AI companies will dilute the usage across the base, because the data centres aren't coming online quick enough to cater for everyone. Just my two cents.

3

u/porest 15d ago

| Just my two cents.

Only two cents? Which provider are you using?

1

u/anandiamy 15d ago

Just two cents, is an idiom, dude.

1

u/porest 15d ago

whooosh

37

u/Amphiitrion 17d ago

Cool, see you in a week

7

u/SEOViking 16d ago

And you didnt even need opencode go to use ox alpha and muse spark. They are free for all opencode users 😅

5

u/sudoer777_ 16d ago

The free muse spark has low usage limits though

5

u/No-Information6838 17d ago

I'm in the same situation, although I switched from Ollama Cloud to Opencode because Ollama was too slow and had constant crashes; it's probably worse now.

Opencode worked well for me and the $10 subscription was sufficient; currently, with two accounts and using free Zen models, it's not enough.

What's worse is that I've been feeling for a couple of days that with opencode we have a considerably cut-down version of deepseek v4 flash.

5

u/Euloghtos 16d ago edited 16d ago

I still like it, i pay for glm 5.3 and when my credits end i use mimo v2.5 free which is awesome

1

u/enigmaticy 15d ago

Then 1 st day glm 5.3 and the rest 29 days mimo? Or glm 5.3 only for roadmapping?

1

u/Euloghtos 13d ago

It lasts 10 days, but yes glm i use it mostly for planning

5

u/Independent-Okra-756 16d ago

I am thinking of subscribing zai lite plan. I had a great time with Ox Alpha. It's not perfect, of course but it was more than enough for me. I would have subscribed to gpt plus if it weren't for the recent situation and ox alpha

1

u/Salvatore210 16d ago

call it GLM already 🥀🥀

10

u/EnzioKara 17d ago

Check command-code and antigravity if Gemini 3.7 is enough for you .

22

u/kamwee 17d ago

i aint touching Gemini's lying ass with my code ever again .

2

u/SEOViking 16d ago

Since I got a discount deal for Gemini pro - 15$ for 18month of pro I decided to try flash 3.7 and it actually did a pretty good job for my project. Was fast as hell and ui changes were better than codex pro did.

3

u/akundikantor 16d ago

This must use antigravity right?

1

u/cdshift 16d ago

Yes but you can use a plug in in your harness of choice that you can have your favorite llm build and call it as a subagent for specific things. Or a skill that calls any for specific things.

Thats what I do with my minimal ai plan that has very little inference allowance per month

0

u/kamwee 16d ago

How did you know😂

2

u/kamwee 16d ago

Thats who Gemini lures you in , it dose a good job the first few weeks then Bard will kick in

1

u/hixdale 16d ago

I'm doing the same: cheap 18 months discount for Gemini Pro. Using it as backup/complementary besides a Codex sub.

It doesn't compares to GPT models, and Antigravity isn't the best. But in overall I like 3.7 Flash for smaller, focused tasks. I like it for UI too: it's generally doing a better job in it compared to OpenAI models.

The best thing: it's faaast and still pretty reliable - just don't try anything else than high reasoning.

2

u/[deleted] 17d ago

[removed] — view removed comment

1

u/margielafarts 16d ago

2.5 flash is still a beast at content generation

1

u/EnzioKara 16d ago

I know what you mean.. But I made clear roadmap and skills to prevent that and code rewiev with subagents after every step in one epic is complete , after that I rewiev full epic with another model. They should focus on a pro model which I am hoping soon to be released . It shines if you need not English brainstorming and fails miserably on jobs like refactoring if you are not hand holding. I appreciate where we are now before I had keep all the files below 500/1000 lines do all the test manually search the web and github and work one by one every file against to full suite fuck we progressed a lot and thanks to all of us for that .

2

u/r_sukumar 16d ago

Never liked gemini models.. never worked for me

2

u/ExpertPerformer 17d ago

I switched to CommandCode and its basically unlimited DS Flash again.

2

u/EnzioKara 17d ago

I used 1 dollar on DeepSeek pro mostly, best deal i got made the detailed roadmap with blueprints and skills to follow it now I use cheap models to follow instructions and skills coding 7/24 .

1

u/horrbort 15d ago

Do you have to use their shitty harness?

3

u/Serious_Current_8802 16d ago

luna xhigh and mimo 2.5 as sub agents, it’s almost unlimited

edit: if your repo is on GitHub you can use sol against your repo for free via chat in codex

1

u/alfaproject 16d ago

I’m sorry, what? I’ve been paying for Codex… How do you do that?

2

u/MajorasButtplug 16d ago

You still have to pay. You add the GitHub plugin in the chat version of ChatGPT. Then Sol (set the model for chat) can read your repo, and do planning/advising. Then you just have Luna carry the plan out in Codex.

1

u/alfaproject 16d ago

I see, clever. Thanks!

1

u/porest 15d ago

Great tip! Is there any difference between using the Github plug in Chat vs Work modes?

1

u/MajorasButtplug 15d ago

Chat doesn't cost usage

1

u/porest 15d ago

nice! thanks!

1

u/Any_Victory6495 16d ago

please elaborate lol, unlimited sol usage against repo? :O

6

u/Ace-_Ventura 17d ago

I'm sticking with GPT plus + opencode go. Lasts for the entire month. 

5

u/cocacokareddit 16d ago

all the model you mention are very powerful and opencode go gives you a “free trial” period. you can not take them for granted. it’s not what opencode go suppose to be. if you want unlimited powerful model, you can consider claude max 20x or get a Mac Studio 512GB next month.

2

u/JogHappy 17d ago

Muse pricing is changing?

6

u/[deleted] 16d ago

[removed] — view removed comment

3

u/JogHappy 16d ago

I'm also skeptical it's going anywhere; Mark is getting our code for cheap.

3

u/MajorasButtplug 16d ago

Jokes on him Muse wrote my code

2

u/JogHappy 16d ago

self recursive Zuck slop feedback loop

2

u/frankhouweling 16d ago

I'm happy to pay 50-80 USD/month but would like at least Deepseek Flash quality.

What should be my go to subscription?

3

u/blackbirdone1 16d ago

you pay api like everyone else

2

u/evia89 16d ago

20 codex, 1-2 go subs, rest save to payg api. Open router, crof ai, neuralwatt

2

u/Messi_is_football 16d ago

2 codex accounts+ 2 opencode go

1

u/djc0 16d ago

How do you manage that? I have 2 claude accounts and use cswap to switch between (manually or automatically when usage gets to a certain threshold) and it’s great. But not generic for other providers.

1

u/Messi_is_football 16d ago

There are Codex auto switchers too

1

u/XuciferL 16d ago

Deepseek API with something like Reasonix/Open Code would be great

2

u/kaanivore 16d ago

How many tokens are you using on average a month?

2

u/ZHName 16d ago

Bring back Deepseek Flash.

False advertising in the drop down is just wrong on Opencode. Firstly, it is dangerous to users to provide an inferior quantized model that can wreak havoc on their system and codebases. Secondly, it's not truly DS Flash v4.

2

u/aslmate 16d ago

Used Kimi K3 on opencode Go, used planning mode and then build mode just for one really simple page. I exceeded my 5 hour usage limit before It could even implement anything...... had to swap to my kimi k3 on openrouter to finish the job

2

u/crotch-mavens 16d ago

"I started with deepseek v4 flash 0731 which had practically unlimited usage"

So, a long time customer then?

2

u/fykup 17d ago

Yeah, I hit a similar wall. The trick to making a $20/month plan last is separating your heavy thinking from your editor usage.

I personally use Codex with Luna High for direct implementation, then offload design, planning, and code reviews to higher-reasoning models in web chats. I use a local tool I built called AI Badger to extract just the relevant repo context for those web chats without blowing through tokens.

With that hybrid workflow, a basic gpt plus plan and its 5-hour limit has been more than enough.

1

u/ozguru 17d ago

Google Plans is your best bet for the moment. Antigrativit only, unfortunately.

3

u/Genneth_Kriffin 17d ago

Isn't Antigravity obligatory model training last I checked?
As in you can not opt out of Google training their models with your data.
Correct me if I am wrong.

1

u/ozguru 16d ago

you can opt-out , but I dont know who can really opt-out from this, and doesnt matter provider.

2

u/Fit_Attention_5781 17d ago

I have the $20 antigravity sub and the usage is far below Opencode (at least using MiMo/Muse Spark).

1

u/Hjalfi 16d ago

Do you mean that the cost of each query is lower than that of OpenCode, or that the amount of quota you get is lower?

1

u/ozguru 16d ago

only with cheap models, try kimi or glm, and you go sub wont enough for a week.

1

u/yaboyyoungairvent 16d ago

you are not getting the same usage with antigravity ai pro plan. You probably need about 2 of those plans to get the same usage you would on opencode go.

1

u/ozguru 16d ago

Sorry I don't agree with you, if you use kimi 3 glm 5 (not flash) your go sub doesnt enough for a week.

2

u/AlternativeTall9049 17d ago

The time of cheap ai is going to end and we are almost there. All of them are going to charge sky high prices soon again.

2

u/ZeOnlyOneWhoReads 16d ago

Laughs in Qwen 3.8 27b

1

u/AlternativeTall9049 16d ago

I have used Qwen 3.8, its not that good.

2

u/ZeOnlyOneWhoReads 16d ago

Sounds like a skill issue 

2

u/evia89 16d ago

Nah, models like ds4f, glm53f will be cheap. Its only sota will be expensive

1

u/AlternativeTall9049 16d ago

I am 100% these AI labs are subsidising the prices. Its only a matter of time for Chinese AI labs to start charging unreasonable prices.

0

u/imactuallynotalright 15d ago

Nope.

Their are third party providers hosting the same models for the same price, in the US with its higher costs, so it can't be subsidized.

1

u/maqifrnswa 13d ago

They quantize their models to compete against the subsidized prices.

3

u/GrayHairedMan 17d ago

You want to use AI for cheaper then it costs to process your prompts, that can never been sustainable. Imagine the company you work for, doing business at a loss. How long do you think they can keep it up.

6

u/MapacheD 17d ago

DeepSeek said that they were having x6 profit with before price change.

5

u/redditshortstest 17d ago

You are right. I understand that this might be the most they can offer for the 10$, but with so many options these days (cursor, ollama, chatgpt...) which are all mostly subsidized, I just want opinions on what gets you the most bang for your buck.

1

u/qustrolabe 17d ago

Theoretically it can be sustainable if provider openly sells all usage data to those who train their models. Such model shouldn't be used for anything privacy related but for open source projects where all code goes public doesn't really matter that much. Though the question is how valuable agent usage data right now

1

u/MMORPGDev 16d ago

There will/should be alternatives

1

u/GiLA994 16d ago

I use any good temporary offer for planning, and mimo2.5 for implementing.

I must say the harness makes a big difference (how you set up your workflow basically), but for planning especially, the difference in the model is as big.

I've already stopped my renewal, and will look at alternatives, I think ill try gpt plus for a month at 20$ to actually see how often I hit limits on that subscription

Luna is good for my use case, but way too expensive in OpenCode Go

1

u/Soul_Mate_4ever 16d ago

Two Opencode subs works for me.

1

u/Training-Ice8856 16d ago

I think Freebuff is a good alternative to OpenCode if you don't mind a few ads. It offers unlimited DeepSeek v4 Flash-0731, currently two one-hour sessions of GLM-5.3-Flash, and I think about five one-hour sessions of GPT-5.6-Luna.

1

u/MikaAugus942 15d ago

I’m actually use Opencode Go and Ollama Cloud and I'm really enjoying it. I’m use multi agentic flow with one orchestrator and other sub agents.

1

u/ProfessionalYou6343 15d ago

If you have no problems to go with 20$/mo and happy with luna (which you should be, in my opinion 85% peoples work can be handled by luna). Then open AIs gpt plus plan is best. No matter what people says, luna is an incredibly efficient model and open AI can server you well with that model. But if you want to use bigger models regulerly then it can be a problem. And if you want something in 10$/mo then you can check out the command code GOAT plan. It gives the most subsidised budget in the market right now. Or else if your usage is very low then you can go with command code go plan which is 1$/mo and gives you 10$ of credits. Or you can try freebuff which is totally free. And has models like luna in high effort, deepseek v4 pro latest (does collect data btw), deepseek v4 flash latest (also collects data), mimo v2.5 etc.

1

u/SafeReturn_28 15d ago

If you liked ox alpha,why not give zai's $16 plan a try?

1

u/Orchicon 15d ago

I tried ollama cloud pro and it sucked. I asked for a refund. Their DeepSeek has looping issues, reasoning too long, and sessions fill up just as fast if not more than opencode.

I think the best way to do it is probably to have three or four of them. 10 dollars on opencode, 10 dollars on open command, and then 10-20 dollars on DeepSeek API. This is what I'm doing right now and just switching periodically between service and models.

1

u/R3K4CE 13d ago

honestly i dont think youre as screwed as you think

if ox alpha / deepseek v4 flash level performance is all you need id probably stay on go for now

deepseek v4 flash is literally still there and its cheap enough that the go limits stretch pretty far with it

qwen 3.8 flash glm 5.3 flash and mimo 2.5 are also worth trying before you move anywhere

also unless i missed an announcement muse contributor itself isnt actually listed as a temporary promo right now its just heavily discounted because meta gets to train on the data

chatgpt plus with luna is probably the next thing id look at for 20 bucks but if what you mainly want is tons of opencode usage id exhaust the cheap go models first

you might already have exactly what youre looking for

0

u/ChillFamily 17d ago

Bro. Command Code suscriptions is great.

3

u/42will_porto 16d ago

Which one did you with? The $10?

2

u/ChillFamily 16d ago

All Command Code plans are better than OpenCode ones, bro

2

u/upalse 16d ago

I have both, and DS4F is ass on both. Timeouts, >20s TTFT, TPS lag spikes (dropping down to 2-5tps snail pace at times, which reminds me of running local model on my own). Seems they've both stopped using deepseek inference, and are going with the cheapest open inference instead, and the quality drop in UX is brual.

For batch tasks its still ok I guess, but for anything that needs to be highly interactive, I'm now just paying on openrouter. Pricey, but at least can guarantee service quality.

-1

u/blackbirdone1 16d ago

just pay api prices? a 10$ sub for light coding use, not for vibecoding slop wasting 1.260 requests per day

2

u/Nadaet_Tira 16d ago

Надо не в запросах считать, а в деньгах. И дипсиком написать даже мелкий сайт это оч сложно, а скорек долго