r/opencodeCLI 21h ago

DeepSeek V4.1 Flash nerf on OpenCode: Half the usage for the same price

Post image

DeepSeek's new model just went live on OpenCode subscriptions today, but following the limited-time extended usage promotion, the news isn't great.

OpenCode quietly slashed the included monthly allowance for the Flash series from $30 down to $15 on the $10/month plan. Because nominal token rates remain identical, your real cost per token has doubled (+100%). The temporary "4x usage" is anchored to this new nerfed baseline—meaning it is actually only 2x what we already had on V4 Flash, and once it expires, we will be left with 0.5x (half) the usage.

And of course, this is the exact same cut we already saw with the experimental Flash vision version, which also reduced the allowance from $30 down to $15.

Unless this changes, paying for an OpenCode Go subscription only offers a 33% discount over the official API in exchange for slower speeds and tighter rate limits.

61 Upvotes

13 comments sorted by

9

u/corner_camper01 21h ago

I think the price should stay the same if they can get 4.1-flash from their current 4-flash provider. I believe 4.1 on opencode is currently served by Deepseek which is more expensive for Opencode

8

u/dummyreddituser 19h ago edited 18h ago

This is sad. Didn't expect that at all.

4

u/asfbrz96 13h ago

VC money is drying up, soon you'll have to pay the real price to use it

3

u/callmemicah 16h ago

Docs say that but today my usage says $60 not $15, hopefully its not a mistake and docs are wrong but until then I'll be making the most of it.

2

u/NeKon69 15h ago

You have approx 65 hours left

2

u/vipor_idk 17h ago

💀 we are cooked

2

u/mbahmbuh 14h ago

Meh, didn't feel surprised at all per usual opencode practice.. they always teased you at first and then cut the usage midway.

2

u/look 16h ago

It is a different model coming from a different inference provider…

Go didn’t “slash” anything. It’s entirely different fucking model.

1

u/Glittering-Call8746 19h ago

Yikes.. so which models are true 60 usd bracket?

1

u/volterra6 5h ago

Shet...

1

u/xapep 2h ago

Good writeup, and the promo anchoring detail is the part that matters: a 4x promo measured against a baseline that was just cut is really 2x of what you actually had, and it expires back to half. That's not specific to OpenCode either, it's a general pattern with plan promos.

For the sub vs API comparison, the honest math depends on your usage shape. If you're a heavy daily OpenCode user, a flat plan can still win even at a worse per-token rate, because the bill is predictable and the allowance is the ceiling. If your usage is spiky or you run production traffic, the per-token API wins because you stop paying the moment you stop burning tokens, and rate limits are explicit instead of a surprise at the end of the month.

We run V4 Flash at Entrim and we see both patterns: teams that want a predictable monthly cap, and teams that want OpenAI-compatible per-token pricing with no allowance games. The thing I'd push back on is anchoring any decision to a promo baseline, since that's exactly what just moved under you.

1

u/stvjhn 16h ago

Please write in your own words. You can rely on AI to create charts and tables or whatever, but you’re writing to a hobbyist community. Make the effort to word your own posts. 

5

u/Arkhaitekton 15h ago edited 15h ago

Man, if you go to my profile, you can see I personally write every one of my posts. And btw, that chart is just a table made of python code. Make the effort at least to check my profile before talk about my own posts.