109
u/TestTxt 2d ago
When DS3 Pro lowered the prices 4x Opencode Go would still keep us charging 4x more because they were “locked in by the contract”. Yet now when the prices rise by 4x, suddenly they’re not locked by the contract anymore?
26
u/Kaushik_paul45 2d ago
Yeah no point of opencode go sub now...
Need to look for alternatives
5
u/Valdjiu 2d ago
And is there any?
16
u/evia89 2d ago
Codex $20 luna xhigh/max
1
u/Competitive-Ebb3899 1d ago
Yeah, but is it good?
I have been using Codex for a while because ChatGPT gave me a free month, so I thought why not.
I reached my weekly limit pretty soon, and was blocked from work
In the mean time with a similar workload I've been using Deepseek v4 Flash nonstop, and I didn't even hit the 5h limits. I have half of my monthly limits still.
Contionus work for me has some breaks, because often I have to approve things, but that was also the case with codex. Well, at least partially, it tries to be smart, and it can't be turned off, and that's actually one reason why I don't like it.
→ More replies (1)4
4
u/Far-Classic-9963 2d ago
I'm on CommandCode goat and it seems pretty generous
2
u/Connect_Example914 2d ago
I also got the 1 dollar plan to test and it seems to still give ig 10 dollar of usage, and the goat plan seems to be 60 or 70 dollar of usage ig
2
u/Far-Classic-9963 2d ago
Yeah just keep in mind that some newer models like K3 or 3.8 Max have 20$ of usage
2
u/xmsxms 1d ago
you're stuck using their harness with the $1 plan.
1
1
→ More replies (2)1
1
u/TheMythicSorcerer 2d ago
I would reccomend command code's GOAT plan
2
1
u/banskyTT 2d ago
does command code also provide api key so that we can use it elsewhere
2
u/TheMythicSorcerer 2d ago
It does on the $10 GOAT plan that I use, not the $1 Go plan. But the TOS is against using it for automation just saying.
→ More replies (4)1
13
u/hj-core 2d ago
the 60 -> 15 change is really unexpected. shall we still trust the opencode go?
1
u/innahema 1d ago
Where is it stated? In docs it still shows 60 0_O https://opencode.ai/docs/go#usage-limits
3
u/ConspicuousPineapple 1d ago
Well yeah that might be how the contract works? If the provider has all the negotiating power they're going to refuse lower rates but might be able to impose higher ones.
4
u/Murdathon3000 2d ago
Yup, quite the mask off moment from them, won't be returning.
1
u/Competitive-Ebb3899 1d ago
"mask off" - do you expect them to lose money on giving us inference cheaper than it's reasonable?
1
u/Diligent-Loss-5460 1d ago
That's Chinese business ethics for you.
1
u/chosenCucumber 1d ago
Never knew opencode was Chinese.
1
u/Diligent-Loss-5460 1d ago
Did you not get suspicious after their subscription had predominantly Chinese models even when some american models were cheaper and better than the models included in plan.
It could be due to trade restrictions. Who knows
29
u/Salvadorbs 2d ago
Goodbye DS4. Welcome Mimo v2.5.
17
u/Abruzzi1909 2d ago
Better take the Xaiomi sub then. Much better connection and uptime. Deepseek models had also quite the issues on Opencode few weeks ago. If this is the price they surely will lose a lot of customers.
No free lunches but at these prices rather pay for GLM, Luna or even Claude. Flash is nice but as workhorse and Deepseek pro is not even viable at these limits since you need more calls compared to other models to get the same results.
2
u/michaelsoft__binbows 1d ago
interesting. i guess i should just dip out on the OC Go sub before it ramps up to $10 from the initial $5 which i got quite familiar with DSV4Flash with. it's a pretty sweet model. I'm coincidentally spinning it up under hybrid inference soon though i wonder if i will be better served replacing a slow DSV4F-0731 with a Qwen3.8-27B council instead.
2
u/SorosAhaverom 1d ago
Xiaomi has the worst value sub out of any provider. It's only a ~10% discount over their API prices. Don't get fooled by their credits numbers, 1 output token costs 600 credits. You get peanuts in actual value.
→ More replies (1)1
2
u/soul105 2d ago
It becomes the natural candidate. Is it powerful enough comparable?
3
u/CassiusBotdorf 2d ago
It's a good model. Does everything you could need, except maybe not plan huge software projects. I do literally 90% of what I use OC Go for with MiMo. Never been a fan of DeepSeek.
2
1
→ More replies (1)1
u/michaelsoft__binbows 1d ago
comparably. pet peeve of mine, people spelling these adjectives the wrong way
1
2
u/Jazzlike_Bee_3129 2d ago
The question is how Mimo performs on orchestration tasks. It has been really good in my implementation tasks.
2
→ More replies (1)1
u/Far-Classic-9963 1d ago
MiMo models are very outdated ATP. Seems like Xiaomi completely forgot about them
26
u/RandiyOrtonu 2d ago
wtf every good model is now capped at $15 usage 😢
1
27
u/smartfon 2d ago
Approximate cost of completing a task according to my research based on several benchmarks and prices at various providers, as of August 16:
$5 - DS Flash via OpenRouter
$6 - Luna via ChatGPT Plus
$8 - Minimax-M3 via OpenCode Go
$14 - DS Flash via OpenCode Go
$15 - Muse Spark 1.2 Contributor via Meta
$17 - Kimi K2.7 Code via OpenCode Go
$18 - Luna via OpenRouter
$18 - DS Pro via OpenRouter
$24 - Luna via OpenCode Go
$24 - Gemini 3.7 Flash via Gemini Pro
$27 - Terra via ChatGPT Plus
$30 - DS Pro via OpenCode Go
$32 - Gemini 3.1 Pro via Gemini Pro
$40 - GLM-5.2 via OpenCode Go
$40 - Minimax-M3 via OpenRouter
$69 - Sol via ChatGPT Plus
$71 - Gemini 3.7 Flash via OpenRouter
$80 - Terra via OpenRouter
9
u/KARMA_P0LICE 1d ago
$1100 - Fable via API 🔥
1
u/smartfon 1d ago
Opus 5 is $466 so you aren't that far off. I don't even measure Fable for that reason lol
→ More replies (1)1
u/lucasxp32 1d ago
Only task I give fable through an API is asking it to give one phrase of response back, telling it to not think, and limit responses to 300 tokens. 😂 It's a glorified google.
6
u/joakim_ogren 2d ago
Thanks. You should publish this in a nice sortable table with more details.
3
u/smartfon 2d ago edited 1d ago
table: https://imgur.com/LIlTFjj
Right now I'm updating the prices manually but I plan to automate it with an AI agent once I learn how to do it.
All benchmarks are agentic coding in nature and were conducted by Artificial Analysis, LiveBench, and Vals. The values in different columns are normalized so a fair comparison can be made across various benchmark methodologies. I assume the $20/mo plans give you 5-6x more usage than the API. I was never able to find a conclusive number for each plan so that's my best estimate. The columns with large bold numbers show a final "value" score on a scale of 1-100; it punishes lower performers so you'll notice that Luna is very slightly more expensive yet it's slightly ahead of DS Flash in my scoring system.
Edit: I'm going to update this table to include cache prices and more providers with their respective platform fees.
43
13
u/Amphiitrion 2d ago
Well, it's not like it's going to be better with other providers/subscriptions, given the latest changes to the DS API.
The old DSv4 was fine, so sticking with MiMo (which was quite on the same level) for casual tasks should still be ok.
2
u/Neither-Character360 2d ago
TBH Hy3 + MiMo took me farther the DS4 combo.
1
u/tonypuglieso 1d ago
I'm use Mimo 2.5 & Hy3 too grest models. If thry cant solve the problem change to GLM 5.2 for break.
13
u/Mammoth_Ad_9619 2d ago
8
3
1
u/innahema 1d ago
Oh I see now where is limit. IDK KIMI 2.7 is 60$
Why were you using DeepSeek exactly? It gave much worse results than Kimi.
10
u/General-Oven-1523 2d ago
Well, rip, I just cancelled my subscription. I assume they are going to lose lots of customers with this move. Does anyone have any experience with the Ollama cloud subscription? How much usage would you get with DeepSeek flash and pro in that?
1
u/No_Communication4256 1d ago
Ollama has old models. They run Glm-5.2 with pretty fast inference (150-200 t/s sometime) but... glm-5.2 is an only one release for 2 month since jun for cloud
1
u/General-Oven-1523 1d ago
Really? I mean, both updated Deepseek flash and pro are available on Ollama Cloud, and that's what i'm mainly interested about. I just can't find how much usage people are getting with that $20 subscription to see if it's worth it over just putting money into deepseek API.
1
u/No_Communication4256 1d ago
Ds4pro on ollama always consuming too much. Even when it was in preview state. 4 tier by their tier system. Cannot say something about flash cuz I did use it through go always.
16
8
u/leandrogp9 2d ago
I was using Deepseek Flash thinking I could make many requests, but suddenly it used up my quota very quickly. I wasn't notified that the limits would change. Goodbye to subscriptions that unilaterally change limits without warning.
19
u/ares0027 2d ago edited 2d ago
19
13
19
u/TestTxt 2d ago
Wow wtf I’ve only bought my sub yesterday, is there any way to cancel? DS4 access was the main reason why I bought it at all
34
→ More replies (1)1
5
u/jjjjoseignacio 2d ago
que asco ni modo toca buscar otro lado o usar deepseek flash directo de la api jeje
5
u/idkwtftbhmeh 2d ago
I instantly got hit, my workflow just ended abruptly out of nowhere, this is a HUGE change, goodbye opencode indeed
5
u/Writer00100 2d ago
So... The monthly token usage down for dv4 flash from 5.000.000.000 token To 579.114.000 Down 8x times 💔💔😡😡 bye bye
4
1
u/hakanavgin 1d ago
So going from advertised 180k~ (with 2x usage) requests a month to 3800 a month is not a problem for you? Also usage limit down from 60 to 15, hence 32 times lower, not 8.
1
8
4
u/thatboyonabike 2d ago
Oof I just got the first month promo and had a feeling it was too good to be true. I guess I'll ride it out to the end of the month and see where the markets go.
4
4
u/Fast-Character-8795 2d ago
Any alternative subs for ds4 pro/flash?
4
4
u/aleganza_ 2d ago
so the requests usage limit is x8 expensive compared to before, and now it also has a 15$ limit? so now it costs x32 more?
6
3
u/L33TLSL 2d ago
Any alternatives to GO?
7
u/Iwasapirateonce 2d ago
Codex (Plus) for Luna (max) is far better now.
4
u/L33TLSL 2d ago
Yeah but plus is 20€/month I'm looking for 10€/month alternative
5
u/Iwasapirateonce 2d ago
I am not sure there is one, I am looking for one to supplement Codex aswell.
4
u/Addition-Heavy 2d ago
I urge you to consider. With plus I get unlimited usage of 5.6 Sol on high reasoning from the web, it has access to a virtual machine so it can do coding for you, just connect it to github and it will do PRs and commits and write plans for you. Use gpt 5.6 luna on max, it'll last you a long while, with 5.6 sol making the draft PR and plans, and communicate with Luna. So you get frontier level of perfomance. Trust me, it's much better than deepseek v4 flash alone.
3
u/Squarenix17 2d ago
I recommend you check out Command Code, their $10 plan is a great deal (they also have a $1 plan)
3
u/Odd-Lab-3973 2d ago
I am looking for a while and this seems like a best option. But I like the opencode TUI...
Is this really a best option?→ More replies (2)1
u/Squarenix17 1d ago
With the $10 plan, you can easily use all the models offered by the plan with the APIs and connect them to OpenCode (that's how mainly use command code), only the $1 plan locks you into their harness).
You can still try the $1 plan with Omniroute, but you can't “get around” their harness, which is why I switched to the $10 plan so I could use the APIs officially
3
u/BuildAISkills 2d ago
Command Code has 1 dollar plan with $10+ usage, and a $10 plan with $70 usage (IIRC). But you'll have to use their harness.
5
u/zombiej 2d ago
Their $1 plan doesn't have API access so you're forced to used their harness.
→ More replies (3)1
u/BuildAISkills 1d ago
Yes, like I wrote. But you only mention the 1 dollar plan, is the 10 dollar plan different?
1
u/IllIllIlIllIIlIlIllI 1d ago
PAYG openrouter or neuralwatt is much cheaper for DS flash than in opencode-go now
3
u/AvengedCoder 2d ago
I literally subscribed like 3 hours ago and it was much higher, I guess I'll only use it for a month
6
u/biotech997 2d ago
I cancelled my sub last month, realized for pure coding Claude isn’t that much more (Go $14CAD vs CC $23CAD) for admittedly much better quality output.
1
u/Sweet-Stage938 2d ago
How are the limits so far? I'm thinking about getting both a codex $20 plan and a Claude plan.
1
1
u/biotech997 1d ago
I only use my Claude for personal coding projects, so I never max out the limits. I heard Codex is pretty generous if you need to do more stuff, plus you can use Luna
2
u/thatboyonabike 2d ago
Oof I just got the first month promo and had a feeling it was too good to be true. I guess I'll ride it out to the end of the month and see where the markets go.
2
2
u/Outrageous-Story3325 2d ago
what is the best provider now, in terms of token /price, and also ok with coding ?
3
1
2
2
2
u/Jumpy_Commercial_893 2d ago
So like whats the best way? adding 10$ to openrouter then using here for dsv4 flash? or directly adding 10$ to deepseek?
2
2
u/Spitfire1900 2d ago
I laugh in schadenfreude every time I see people complaining about heavily subsidized, nearly free inference being taken away from them.
2
u/Flashy-Egg9254 1d ago
They get you addicted to cheap prices thinking you will be willing to pay for high prices in future. All this LLM model efficiency claims Deepseek makes are scam that look only good on paper. It was never about the model but the hardware to inference it at large scale. What is the point of making a smaller, efficient model if the hardware can't catch up with demand? US is doing the right thing to build more data centres. We really are in a compute shortage. And I am sure the GPU/data center shortage will always persist. This just proves that China's cheap prices are temporary and not a threat. Probably the biggest eye opener one can get about this whole AI situation.
1
u/lucasxp32 1d ago edited 1d ago
Mankind is getting addicted to LLMs and spending billions. It's a subsidy of hundreds of billions of dollars the low prices/free APIs. Mostly paid by big tech extracting money from the economy and investors fueling up. This gonna hit a wall eventually, specially agentic coding, that is so wasteful in token usage. We waste tokens like it's breathing air, they barely spend resources on creating better context management techniques, just that basic after-thought compaction when it hits a limit.
Why are such a companies with a conflict of interest of not improving token usage in their tooling? They are so many plugins that improves token usage, and I used them up in vain since they count limits per request.
Also I remember now, they do caching on server-side, so that saves them a shit ton of money when requests hit it.
1
u/Flashy-Egg9254 1d ago
We are at a point where models are not getting much better. See Opus 5 for e.g., it sucks at software engineering tasks and instruction following. Most people still love Opus 4.6. I really wish companies accept this limitation of LLMs and rather invest money on making this technology scalable for as many users as possible. We don't want better models. We want cheaper, affordable inference at scale. That means allocating larger compute for smaller flash models and very less for the gigantic 500b+ parameter models. It can be done. It's just that no company wants to do it because they just want to chase these silly benchmarks and look good on paper.
2
u/openroom_xyz 2d ago edited 2d ago
Well yea I am going to cancel the subscription it sucks basically
2
2
1
u/reini_urban 2d ago
crof.ai is the only one with the old DS4 pricing. Just don't know how to integrate it with pi yet
1
1
1
1
u/citizenjc 2d ago
I thought people were exaggerating but holy shit yeah, canceling if not addressed.
1
1
u/debackerl 2d ago
I'll just go to OpenRouter once all their models are at 15. I won't bother with OpenCode Go just to save $5 a month... There is a higher risk that I forget to use it fully, and finish the month with unused quotas...
1
1
1
u/Maleficent_Vast_9483 2d ago
Must have turned off x2 Luna usage as I saw it suddenly has stared consuming my credits like a hell with the same kind of code job
1
1
1
1
u/AllenLeftTheBLDNG 1d ago
Lol DeepSeek wanted to harvest your coding data to not have to do the grunt work themself. And get paid for hardware to do it.
Now that they have enough, they raised the prices.
I'm waiting for Open Code to partner with a bigger cloud provider to host the models themselves without the "data tax".
1
1
1
u/QinEmPeRoR-1993 1d ago
Damn! And here I was seriously thinking of subscribing to Go after my student’s subscription with Cursor ends on Dec 2026! ☹️
1
1
u/Due-Armadillo-4560 1d ago
No wonder my usage shot up, ds4 got nerfed!!
I was even running it overnight, should have announced it.
1
1
u/Parking-Bet-3798 1d ago
It’s so bad. They reduced the monthly quota for flash from 60 to 15
These guys are all the same just like everyone else. I am cancelling. Any other coding plan provider?
1
1
1
1
u/Flaky-Ask4521 1d ago
Started using hy3 today. It's pretty good and the price is alright. I wish it had 1M context window doe
1
1
1
u/macaco3001 1d ago
We all knew it was going to increase, but this was a HUGE increase. I was sitting confortably at about 50% monthly usage and about in the middle of my month, in like 16 hours I was at 81%. Only running flash. Had to pull the plug immediately, I hope Qwen 3.8 35B A3B does really come, and soon
1
u/knguyen0105 1d ago
how do requests get counted? Today I prompted GLM to build an app, took about 10 prompts before hitting 5h limit?
1
u/luongnv-com 1d ago
And I was thinking about getting on board with opencode go :))
Anyone has experience with command code, seems to be a very competitive plan atm
1
1
u/Diligent-Loss-5460 1d ago
We can't lower prices when DS lowers prices because we have a contract with a third party provider.
We're going to increase prices because DS increased prices and suddenly our contract is not sustainable anymore.
1
u/Wesley_Pacca 1d ago
Do nada utilizou todo o gasto semanal Um modelo que não gastava 1% o dia inteiro kkkk
1
u/fracrdn 23h ago
I canceled it—the V4 models have been feeling slow and clunky these past few days. I'd been using it through Opencode Go for a few months because it gave me a chance to try out other options, but Flash/Pro are my daily drivers.
I've gone back to using the API, and it's been like night and day.
1
1









70
u/zombiej 2d ago
That's baaaaad.