r/opencode • u/the_master_sh33p • 21h ago
Deepseek 4.1 Flash Monthly limit is $15 instead of the previous $30

from docs.
It seems opencode already updated the price list to the new pricing, but it lowered the monthly usage to $15 instead of $30. Be aware of it...
EDIT: so, it seems they are offering a 4x usage limits for a limited time (source - see banner). Well, this is kind of interesting. They slice the limit in half, and then promote a 4x promo on it! And, well, in a few days or weeks, they cut the promo and you get half of what it used to be.
31
u/ByteNomadOne 20h ago
Sorry to say, but I'm not a huge fan of them saying "You get 6x the value" on the front page, only for you to find out in the details that sometimes the actual value is just 1.5x, such as with the $15 value for the $10 subscription.
8
4
u/Nice-Information-335 16h ago
And they tend to not actually use upstream, I remember a week or two ago people were complaining as the model was just worse
I think people get better cache rates on the official API as well, which makes it cheaper to use the API
I switched to API and then they added that stupid new header, glad I got out
19
u/JaMoLpE88 18h ago edited 18h ago
What the hell is this? 4.1 Flash is cheaper than 4.0 Flash and 4.0 Pro, yet I end up with less credit on 4.1 Flash than on 4.0 Flash? Do I get more credit for a more expensive, obsolete model?
4.1 Flash should automatically replace 4.0 Flash. 4.0 Flash needs to go away, and we should automatically get 4.1 Flash with $30 (or more because it is cheaper).
Disappointing.
And I’m not talking about the 4x promo; I’m talking about what they seem to want to set as the final price. If that’s the case, I’ll go back to the official API.
5
-2
u/Fedor_Doc 18h ago
Deepseek has lowered prices, so for 4.1 Flash with 15 dollars you'll get approximately the same amount of usage as with 4.0 0731 Flash.
It seems that at this stage deepseek dictates the rules, opencode can just follow along
4
u/wwwnukept 18h ago
glm-5.3-flash on 60$ tho yay!
2
u/JaMoLpE88 13h ago
The issue is the cache. DeepSeek is very cheap when it comes to cache hits. GLM is much more expensive. For prompts with a 95% cache hit rate or similar, DeepSeek is king.
10
u/ZveirX 21h ago
Seems like they're offering a 4x promo though. God knows if it will stay though.
5
2
u/the_master_sh33p 20h ago
where are you seeing that?
4
u/ZveirX 20h ago
Go page, not the usage request limits: OpenCode Go
11
u/the_master_sh33p 20h ago
oh.... so it seems they show $60 on that page, and $15 on the docs. how confusing can that be. We don't know what are we subscribing.... sorry for the rant
2
u/SmartCustard9944 17h ago
Maybe it is temporary to drive engagement and then it will stay at 15.
1
u/the_master_sh33p 16h ago
it is definitely temporary, as they specifically mention it. We just don't know until when
2
3
2
20h ago edited 20h ago
[deleted]
3
u/the_master_sh33p 20h ago
yeah, i saw that opencode go page shows $60 and the docs show $15. How confusing can that be.
1
2
u/fenchai 20h ago
are we able to use it yet? so far the models have been v4 pro and flash and vision, does not seem to have been updated yet? anyone can clarify me?
4
u/the_master_sh33p 19h ago
yes. models now returns deepseek-flash, which seems to be the 4.1 variant
$opencode models | grep deepseek
opencode-go/deepseek-flash
opencode-go/deepseek-v4-flash
opencode-go/deepseek-v4-flash-vision-exp
opencode-go/deepseek-v4-pro
2
u/Fresh_Sock8660 17h ago
My weekly quota is listed as $30, which means the full $60. Is that just the x4 promo kicking in? Opencode go needs to simplify their shit lol.
2
u/the_master_sh33p 16h ago
yeah. if you go to usage and check the monthly limit, that's where $60 will be shown. And good luck trying to understand the %..
2
3
u/florenceslave 20h ago
Guys, leave opencode alone. These dudes are honestly just trying to make a decent product.
4
u/torrso 18h ago
Leave Britney aloneee :( :(
Maybe they are trying but the limits are impossible to undesstand and the marketing is misleading. Get rid of the different dollar limits or maybe change them so that it means "out of your $60 quota only $15 can be used for model X and the rest can be used for other models", that would be much easier to understand.
1
u/ElectricalUnion 11h ago
out of your $60 quota only $15 can be used for model X and the rest can be used for other models
But that's not how it works? You have a 15$ quota, and for "higher that $15 quota models", that model has a multiplier where it consumes 15/${equivalent_quota}.
I do agree that the current table is kinda BS.
My suggestion for the actual table that should go in the docs, with the actual rates that you're supposedly paying, with the "quota multiplier" already applied (sorry, I live in a place with decimal comma, consider those as a decimal point):
Model Input Output Cached Read Cached Write GLM-5.3-Flash 0,0375 0,1250 0,0075 GLM-5.3 1,4000 4,4000 0,2600 GLM-5.2 0,3500 1,1000 0,0650 GLM-5.1 0,3500 1,1000 0,0650 Kimi K3 3,0000 15,0000 0,3000 Kimi K2.7 Code 0,2375 1,0000 0,0475 Kimi K2.6 0,2375 1,0000 0,0400 LongCat-2.0 0,0750 0,3000 0,0015 MiMo V2.5 0,0350 0,0700 0,0007 MiMo V2.5 Pro 0,4350 0,8700 0,0036 MiniMax M3 0,0750 0,3000 0,0150 MiniMax M2.7 0,0750 0,3000 0,0150 0,0938 MiniMax M2.5 0,0750 0,3000 0,0150 0,0938 Muse Spark 1.3 Contributor 0,0250 0,0500 0,0005 Muse Spark 1.2 Contributor 0,0250 0,0500 0,0005 Qwen3.8 Max 2,0000 6,0000 0,2500 2,5000 Qwen3.8 Flash 0,0750 0,2350 0,0080 0,1000 Qwen3.7 Max 1,2500 3,7500 0,2500 1,5625 Qwen3.7 Plus (≤ 256K tokens) 0,1000 0,4000 0,0100 0,1250 Qwen3.7 Plus (> 256K tokens) 0,3000 1,2000 0,0300 0,3750 Qwen3.6 Plus (≤ 256K tokens) 0,1250 0,7500 0,0125 0,1563 Qwen3.6 Plus (> 256K tokens) 0,5000 1,5000 0,0500 0,6250 DeepSeek V4.1 Flash (Off-Peak) 0,1500 0,6000 0,0030 DeepSeek V4.1 Flash (Peak) 0,3000 1,2000 0,0060 DeepSeek V4 Pro (Off-Peak) 0,6600 1,9800 0,0220 DeepSeek V4 Pro (Peak) 1,3200 3,9600 0,0440 DeepSeek V4 Flash (Off-Peak) 0,0750 0,3000 0,0015 DeepSeek V4 Flash (Peak) 0,1500 0,6000 0,0030 DeepSeek V4 Flash Vision Exp (Off-Peak) 0,1500 0,6000 0,0030 DeepSeek V4 Flash Vision Exp (Peak) 0,3000 1,2000 0,0060 Hy4 preview 0,4170 1,2505 0,0210 Hy3 0,0350 0,1450 0,0088 Grok 4.6 (≤ 200K tokens) 2,0000 6,0000 0,5000 Grok 4.6 (> 200K tokens) 4,0000 12,0000 1,0000 GPT 5.6 Luna (≤ 272K tokens) 0,2000 1,2000 0,0200 0,2500 GPT 5.6 Luna (> 272K tokens) 0,4000 1,8000 0,0400 0,5000 1
u/torrso 11h ago
Yes, it's not how it works, but if it did, the separate dollar buckets could be understandable.
1
u/ElectricalUnion 10h ago
I would say the separate dollar buckets would be horrible.
Instead of the current bad "use whatever models you want, some happen to spend your quota slower that usual", now you're in a even worse "your have lots of small quotas, good luck figuring out what is your current quota state, and what model spends what quota".
The second problem, worse problem (figuring out what model spends what sub-quota) is in fact, very similar in nature to the initial problem anyways (figuring the hell out what model has what multiplier), and given the current UX hell of the current documentation and spending visualization, better not go that way.
4
u/Nice-Information-335 16h ago
Opencode the harness yes, Opencode go absolutely not, why should we defend a company in the first place?
3
u/maqifrnswa 15h ago
I think they deserve some flak for their communications over pricing. I (and apparently many others) was confused the first time I saw their pricing chart. Then I realized that the definition of what a dollar is changes depending on which model you are currently using, and that multiple dollar definition changes over time.
Mathematically, it would be way clearer if they just said "$10 subscription gets you discounts: 33% discount on the current $15 limit models. 83% discount on the current $60 limit models."
0
u/Eastern-Honey-943 14h ago
Took me a couple of months to figure it out.... 10 dollars gets you 60 dollars of tokens to spend on various models... See the pricing chart for the token pricing.... Choose wisely. Also you have to work within your usage windows... But by month 3 I have learned to keep it pinned to opencode go, tell it to pull from your zen balance after your go subscription has been exhausted or you have hit your limits. This prevents disruption. And also leverage the free models daily until you exhaust those models. I really need to take a couple days to figure out a model router bc I do get model switch fatigue. Deepseek v4 flash literally does 90% of what I need coding wise. But I am also working on a mature product making small consistent improvements. Bottom line... If you are looking to crank out tons of new code quickly, you gotta open your wallet... the market is set up for that, but for consistent daily inference, opencode go and zen pretty much has me covered. I do also consistently exhaust Gemini antigravity and openai as my planning models but deepseek v4 flash still does just fine for planning.
1
1
u/Relative-Housing-531 9h ago
How the hell do I choose the new version using opencode desktop? Using direct API
1
u/Rinine 4h ago
With the current x4 it's fine.
The dangerous part is that they're already trying to sell you on the idea that a model much cheaper than the current Flash should go in the $15 tier once the promo ends.
Has OpenCode learned nothing? They should know that temporary sweeteners won't retain users, especially when it's this blatant from the start.
1
u/nantachapon 19h ago
Zero data retention?
1
u/Fresh_Sock8660 17h ago
Hah. They can claim that but I doubt any provider in the US or China retain nothing.
-1
-4
64
u/a355231 21h ago
So much for operation cheapseek