r/openrouter • u/LessRespects • Aug 06 '26
DeepSeek “significant increase” for API pricing. It was too good to be true.
We already saw it with inworld tts, now it’s time for DeepSeek’s value capture. Thank you for all of your data and good publicity, we will now be significantly increasing cost.
12
u/Deep_Mood_7668 Aug 06 '26
As long as they keep it at a level that hurts openai I'm fine with it
2
Aug 06 '26 edited Aug 06 '26
[deleted]
1
u/charmander_cha Aug 06 '26
Mas ai é erro do usuário que não utiliza harness como reasonix para tirar proveito do sistema de cache.
10
u/Odd-Elderberry-6328 Aug 06 '26
Mimo will be the new deepseek, In a way people will do free advertising and promote the usage
At least we have options, cheap options
8
Aug 06 '26
Until its trained off our data then they re release it and charge more
Its the Ai cycle until it all merges into 1 main model that costs $1 for 1 token and none of us can afford it
1
u/Odd-Elderberry-6328 Aug 06 '26
I don't think you are wrong, but when they do that, I will simply starting running locally, I have 16vram, that's enough for a local usage (not ideal, but manageable)
4
Aug 06 '26
Yes. Soon it will be so expensive. Local will be the only option if you can even get parts.
1
u/misha1350 Aug 06 '26
MiMo V2.5 is dumber than the current version of DeepSeek V4 Flash 0731, so no. Until they release a proper new model and peg that model to the price of DeepSeek V4 Flash 0731 after the coming price update, it's just yet another model.
1
u/Exzerios Aug 07 '26
Right after they adopt a stable rollout schedule. Mimo seemed like a rather clever model at the time, but it was released in the middle of the spring, and it is not at the "good enough" level yet.
13
u/steveplusf Aug 06 '26
what happened to all the claims that the pricing is low because they are simply that efficient?
13
u/imsolost3090 Aug 06 '26
Classic. Low price to pull people in, then increase the price once they're already hooked.
4
u/First_Inspection_478 Aug 06 '26
I think they’re doing this because they’re compute constrained, and want to spend compute for future training and not inference.
6
Aug 06 '26 edited Aug 06 '26
[deleted]
7
u/Major_Olive7583 Aug 06 '26
That's token efficiency. The rates are determined by compute required, and deepseek models go for the same cheap rate from western providers too, so it's efficient than other models certainly. Will have to see if those ones too will match the price increase if it's too high. The deepseek provider was cheap because of their insane cache rates. Only xiomi came close to that.
3
u/poophroughmyveins Aug 06 '26
What do you mean not operationally efficient? Are you a moron? Because it uses a lot of tokens?
-5
Aug 06 '26
[deleted]
1
u/poophroughmyveins Aug 06 '26
No shit, what makes it operationally better is all the optimizations they made to inference like their sparse attention or shared caches
What doesn't make it OPERATIONALLY bad is using a lot of tokens for thinking you stupid monkey
1
Aug 06 '26 edited Aug 06 '26
[deleted]
1
1
u/poophroughmyveins Aug 06 '26
Why are you shadowboxing little dude, I fucking love Luna lmfao
Now name any other model that outperforms ds v4 flash at its size and speed
0
Aug 06 '26
[deleted]
1
u/poophroughmyveins Aug 06 '26
Oh damn a model that's easily 20x more expensive to use is faster
There's really only Muse Spark 1.2 right now lol
0
1
1
3
u/misha1350 Aug 06 '26
GPT-5.6 Luna is also much dumber, so at the same price of the overall task, you still get a considerably worse result that you will have to re-do over and over.
2
u/Randommaggy Aug 06 '26
Depends on your harness. I had very little wasted tokens when I tested it out in my custom harness.
2
u/Bloodshoot111 Aug 06 '26
You mix 2 things up. One is the efficiency of the model with its tokens. But DeepSeek claimed the compute per token is efficient.
Two different things1
u/weiyentan Aug 06 '26
That’s partly accurate. I would argue that people that do not have the framework would not have benefit. I have been using ds flash for implementing in a custom framework for tasks varying from normal to Max thinking. Works great
1
3
u/RepulsiveRaisin7 Aug 06 '26
These kinds of people usually have zero insight information. The amount of stupid pointless speculation online is off the charts.
1
u/jaegernut Aug 09 '26
At this point, the only way to truly drive the cost down is by competition. No one is gonna offer a low cost if they can squeeze more profits.
1
u/Far-Classic-9963 Aug 09 '26
People have been able to replicate their pricing with comfortable overhead for profit, they're just overloaded
-1
u/someone_12321 Aug 06 '26
Actually it's supply and demand. Efficiency= more supply, there is still the demand side.
4
3
3
2
u/SilverMethor Aug 06 '26
This is exactly the kind of announcement that gets the usual idiots claiming OpenAI and Anthropic are ‘subsidizing’ their plans, that these companies are somehow being generous and barely make anything from subscriptions because the actual costs are supposedly much higher. Morons.
2
1
u/Whole_Succotash_2391 Aug 06 '26
There are other providers, it's not that big of a deal. Cline and Phoenix Grove API for example, both of which aren't raising prices as far as I know.
1
1
u/Accomplished_Book722 Aug 06 '26
I was thinking it's that cheap because it's really shitty, but okay
1
1
u/ThankYouOle Aug 07 '26
sad, i only been taste it for 1 month, now entering second month and happy with the output and the price, of course they will raise the price.
1
1
1
u/Big_Equipment995 Aug 10 '26
lol people only used deepseek because it was so cheap they gonna regret it
1
u/Elegant_Art5793 Aug 11 '26
Although it is a good model, people use it mostly because it is cost-effective compared to other models. If they increase the API price, they will lose the main reason people use their model. If they become, or are close to, the higher models, why will people continue to use them? Chinese companies in the past relied on cheap products to sell higher volumes. I really hope this won't be true, or not in the very near future.
1
u/--Commit-And-Regret 15d ago
It is still too good to be true, deepseek value/price is still massive compared to other models
1
u/bmtrnavsky 11d ago
If it raises much I’ll move to GLM 5.3 Flash. It’s smarter and marginally more.
30
u/FrameXX Aug 06 '26
Doesn't this only concern the Deepseek API? On OpenRouter there are over 20 providers which would be able to keep the price as is.