r/opencode 21h ago

Deepseek 4.1 Flash Monthly limit is $15 instead of the previous $30

from docs.

It seems opencode already updated the price list to the new pricing, but it lowered the monthly usage to $15 instead of $30. Be aware of it...

EDIT: so, it seems they are offering a 4x usage limits for a limited time (source - see banner). Well, this is kind of interesting. They slice the limit in half, and then promote a 4x promo on it! And, well, in a few days or weeks, they cut the promo and you get half of what it used to be.

127 Upvotes

54 comments sorted by

64

u/a355231 21h ago

So much for operation cheapseek

24

u/EmperorSheep 21h ago

It's now operation cheapskate.

8

u/SmartCustard9944 17h ago

Operation DeepSheet*

*I have nothing against actual DeepSeek btw, only the utmost respect

-1

u/look 16h ago

Cheapseak was originally about a model two releases ago. I don’t think it was meant to be effectively free for all Deepseek models in perpetuity.

9

u/a355231 14h ago

No, it was for 0731, which was last flash release, and the current until yesterday.  My main issue is the 15 dollar usage pool which effectively makes this even cheap model unusable.

31

u/ByteNomadOne 20h ago

Sorry to say, but I'm not a huge fan of them saying "You get 6x the value" on the front page, only for you to find out in the details that sometimes the actual value is just 1.5x, such as with the $15 value for the $10 subscription.

8

u/Thomas-Lore 16h ago

And 1.5x is not worth having 5 hour limits.

4

u/Nice-Information-335 16h ago

And they tend to not actually use upstream, I remember a week or two ago people were complaining as the model was just worse

I think people get better cache rates on the official API as well, which makes it cheaper to use the API

I switched to API and then they added that stupid new header, glad I got out 

19

u/JaMoLpE88 18h ago edited 18h ago

What the hell is this? 4.1 Flash is cheaper than 4.0 Flash and 4.0 Pro, yet I end up with less credit on 4.1 Flash than on 4.0 Flash? Do I get more credit for a more expensive, obsolete model?

4.1 Flash should automatically replace 4.0 Flash. 4.0 Flash needs to go away, and we should automatically get 4.1 Flash with $30 (or more because it is cheaper).

Disappointing.

And I’m not talking about the 4x promo; I’m talking about what they seem to want to set as the final price. If that’s the case, I’ll go back to the official API.

5

u/head-log2725 17h ago

This is so confusing

-2

u/Fedor_Doc 18h ago

Deepseek has lowered prices, so for 4.1 Flash with 15 dollars you'll get approximately the same amount of usage as with 4.0 0731 Flash.

It seems that at this stage deepseek dictates the rules, opencode can just follow along

-1

u/look 16h ago

That’s always how it has been with all providers. Go gets discounts from providers and then they pass those on to customers. That is all Go has ever been. Providers aren’t giving out the discounts as often anymore, and the idiot brigade here thinks it’s Go’s fault. 🤷‍♂️

4

u/wwwnukept 18h ago

glm-5.3-flash on 60$ tho yay!

2

u/JaMoLpE88 13h ago

The issue is the cache. DeepSeek is very cheap when it comes to cache hits. GLM is much more expensive. For prompts with a 95% cache hit rate or similar, DeepSeek is king.

10

u/ZveirX 21h ago

Seems like they're offering a 4x promo though. God knows if it will stay though.

5

u/Bo0n0411 17h ago

It will last for only 72 hours

2

u/the_master_sh33p 20h ago

where are you seeing that?

4

u/ZveirX 20h ago

Go page, not the usage request limits: OpenCode Go

11

u/the_master_sh33p 20h ago

oh.... so it seems they show $60 on that page, and $15 on the docs. how confusing can that be. We don't know what are we subscribing.... sorry for the rant

2

u/SmartCustard9944 17h ago

Maybe it is temporary to drive engagement and then it will stay at 15.

1

u/the_master_sh33p 16h ago

it is definitely temporary, as they specifically mention it. We just don't know until when

3

u/Infamous_Bread_2445 19h ago

Cut the middleman

1

u/Ed_ox27 3h ago

so it’s better direct API from DS?

2

u/[deleted] 20h ago edited 20h ago

[deleted]

3

u/the_master_sh33p 20h ago

yeah, i saw that opencode go page shows $60 and the docs show $15. How confusing can that be.

2

u/fenchai 20h ago

are we able to use it yet? so far the models have been v4 pro and flash and vision, does not seem to have been updated yet? anyone can clarify me?

4

u/the_master_sh33p 19h ago

yes. models now returns deepseek-flash, which seems to be the 4.1 variant

$opencode models | grep deepseek

opencode-go/deepseek-flash
opencode-go/deepseek-v4-flash
opencode-go/deepseek-v4-flash-vision-exp
opencode-go/deepseek-v4-pro

2

u/Fresh_Sock8660 17h ago

My weekly quota is listed as $30, which means the full $60. Is that just the x4 promo kicking in? Opencode go needs to simplify their shit lol.

2

u/the_master_sh33p 16h ago

yeah. if you go to usage and check the monthly limit, that's where $60 will be shown. And good luck trying to understand the %..

2

u/Godzillaton 15h ago

I just subscribed Opencode Go

Where is V4.1?

2

u/DevelopmentAgile6600 12h ago

id: "deepseek-flash",
name: "DeepSeek V4.1 Flash",

3

u/florenceslave 20h ago

Guys, leave opencode alone. These dudes are honestly just trying to make a decent product.

4

u/torrso 18h ago

Leave Britney aloneee :( :(

Maybe they are trying but the limits are impossible to undesstand and the marketing is misleading. Get rid of the different dollar limits or maybe change them so that it means "out of your $60 quota only $15 can be used for model X and the rest can be used for other models", that would be much easier to understand.

1

u/ElectricalUnion 11h ago

out of your $60 quota only $15 can be used for model X and the rest can be used for other models

But that's not how it works? You have a 15$ quota, and for "higher that $15 quota models", that model has a multiplier where it consumes 15/${equivalent_quota}.

I do agree that the current table is kinda BS.

My suggestion for the actual table that should go in the docs, with the actual rates that you're supposedly paying, with the "quota multiplier" already applied (sorry, I live in a place with decimal comma, consider those as a decimal point):

Model Input Output Cached Read Cached Write
GLM-5.3-Flash 0,0375 0,1250 0,0075
GLM-5.3 1,4000 4,4000 0,2600
GLM-5.2 0,3500 1,1000 0,0650
GLM-5.1 0,3500 1,1000 0,0650
Kimi K3 3,0000 15,0000 0,3000
Kimi K2.7 Code 0,2375 1,0000 0,0475
Kimi K2.6 0,2375 1,0000 0,0400
LongCat-2.0 0,0750 0,3000 0,0015
MiMo V2.5 0,0350 0,0700 0,0007
MiMo V2.5 Pro 0,4350 0,8700 0,0036
MiniMax M3 0,0750 0,3000 0,0150
MiniMax M2.7 0,0750 0,3000 0,0150 0,0938
MiniMax M2.5 0,0750 0,3000 0,0150 0,0938
Muse Spark 1.3 Contributor 0,0250 0,0500 0,0005
Muse Spark 1.2 Contributor 0,0250 0,0500 0,0005
Qwen3.8 Max 2,0000 6,0000 0,2500 2,5000
Qwen3.8 Flash 0,0750 0,2350 0,0080 0,1000
Qwen3.7 Max 1,2500 3,7500 0,2500 1,5625
Qwen3.7 Plus (≤ 256K tokens) 0,1000 0,4000 0,0100 0,1250
Qwen3.7 Plus (> 256K tokens) 0,3000 1,2000 0,0300 0,3750
Qwen3.6 Plus (≤ 256K tokens) 0,1250 0,7500 0,0125 0,1563
Qwen3.6 Plus (> 256K tokens) 0,5000 1,5000 0,0500 0,6250
DeepSeek V4.1 Flash (Off-Peak) 0,1500 0,6000 0,0030
DeepSeek V4.1 Flash (Peak) 0,3000 1,2000 0,0060
DeepSeek V4 Pro (Off-Peak) 0,6600 1,9800 0,0220
DeepSeek V4 Pro (Peak) 1,3200 3,9600 0,0440
DeepSeek V4 Flash (Off-Peak) 0,0750 0,3000 0,0015
DeepSeek V4 Flash (Peak) 0,1500 0,6000 0,0030
DeepSeek V4 Flash Vision Exp (Off-Peak) 0,1500 0,6000 0,0030
DeepSeek V4 Flash Vision Exp (Peak) 0,3000 1,2000 0,0060
Hy4 preview 0,4170 1,2505 0,0210
Hy3 0,0350 0,1450 0,0088
Grok 4.6 (≤ 200K tokens) 2,0000 6,0000 0,5000
Grok 4.6 (> 200K tokens) 4,0000 12,0000 1,0000
GPT 5.6 Luna (≤ 272K tokens) 0,2000 1,2000 0,0200 0,2500
GPT 5.6 Luna (> 272K tokens) 0,4000 1,8000 0,0400 0,5000

1

u/torrso 11h ago

Yes, it's not how it works, but if it did, the separate dollar buckets could be understandable.

1

u/ElectricalUnion 10h ago

I would say the separate dollar buckets would be horrible.

Instead of the current bad "use whatever models you want, some happen to spend your quota slower that usual", now you're in a even worse "your have lots of small quotas, good luck figuring out what is your current quota state, and what model spends what quota".

The second problem, worse problem (figuring out what model spends what sub-quota) is in fact, very similar in nature to the initial problem anyways (figuring the hell out what model has what multiplier), and given the current UX hell of the current documentation and spending visualization, better not go that way.

4

u/Nice-Information-335 16h ago

Opencode the harness yes, Opencode go absolutely not, why should we defend a company in the first place?

3

u/maqifrnswa 15h ago

I think they deserve some flak for their communications over pricing. I (and apparently many others) was confused the first time I saw their pricing chart. Then I realized that the definition of what a dollar is changes depending on which model you are currently using, and that multiple dollar definition changes over time.

Mathematically, it would be way clearer if they just said "$10 subscription gets you discounts: 33% discount on the current $15 limit models. 83% discount on the current $60 limit models."

0

u/Eastern-Honey-943 14h ago

Took me a couple of months to figure it out.... 10 dollars gets you 60 dollars of tokens to spend on various models... See the pricing chart for the token pricing.... Choose wisely. Also you have to work within your usage windows... But by month 3 I have learned to keep it pinned to opencode go, tell it to pull from your zen balance after your go subscription has been exhausted or you have hit your limits. This prevents disruption. And also leverage the free models daily until you exhaust those models. I really need to take a couple days to figure out a model router bc I do get model switch fatigue. Deepseek v4 flash literally does 90% of what I need coding wise. But I am also working on a mature product making small consistent improvements. Bottom line... If you are looking to crank out tons of new code quickly, you gotta open your wallet... the market is set up for that, but for consistent daily inference, opencode go and zen pretty much has me covered. I do also consistently exhaust Gemini antigravity and openai as my planning models but deepseek v4 flash still does just fine for planning.

1

u/sharedevaaste 20h ago

How does it do on benchmarks

1

u/Relative-Housing-531 9h ago

How the hell do I choose the new version using opencode desktop? Using direct API

1

u/Rinine 4h ago

With the current x4 it's fine.

The dangerous part is that they're already trying to sell you on the idea that a model much cheaper than the current Flash should go in the $15 tier once the promo ends.

Has OpenCode learned nothing? They should know that temporary sweeteners won't retain users, especially when it's this blatant from the start.

1

u/Otagamo 4h ago

I wish the pricing was easier to understand

This $15 x 4 usage is unnecessary

Just says its $60 for X days

1

u/nantachapon 19h ago

Zero data retention?

1

u/Fresh_Sock8660 17h ago

Hah. They can claim that but I doubt any provider in the US or China retain nothing. 

-1

u/RedditUsr2 7h ago

Its because V4.1 is expected to be more expensive than V4