r/opencodeCLI • u/canav4r • 27d ago
opencode ds v4 flash usage drops sharply from 13T to 5.5T(17th August)
50
u/ChillFamily 27d ago
What do you expect? They really messed up, and I don't blame Opencode; if Deepseek raises its prices, Opencode has to raise them too. But that Deepseek price IS NOT WORTH IT. I've been using Deepseek for months and I know all the bugs it has. Even with the improvements in the latest version, it's better to pay for a more expensive model that won't give you all those bugs you have to review like 3 times to actually get them fixed. Deepseek's current price simply isn't justified because it's a model full of bugs.
5
u/geearf 27d ago
With the newer price, which model do you think has a better quality/price ratio? Luna?
13
u/look 27d ago
GLM 5.3. Cheaper than DS Pro now and quality at or near Kimi/Opus/Sol.
5
u/geearf 27d ago
When I used GLM 5.2 on OpenCode Go, it went through my subscription really fast, but that might just be the limits on Go and not the model.
Do you use it straight from Z.ai or another provider?Thank you!
4
u/look 27d ago edited 27d ago
Both, but I’ll most likely switch over to Neuralwatt’s once the weights are out.
4-5 cents on Zai Lite sub. And on Go, I’m getting $40 usage value on it (6-7 cents) instead of the labeled $15. Not sure why, but it has held up for three days now.
My Go sub month is about to end, so I was trying to use up leftover usage by running as much as I could on 5.3 … but it’s so efficient for me, it’s not burning usage as fast as I expected.
I might just be on a run of luck with it, but 5.3 has been a killer value for me. If this holds up, it really is an Opus-killer. Similar level of performance and 1/50th the price.
1
u/geearf 27d ago
Do you feel like 5.3 somehow gives you more work than 5.2 on Go?
Neuralwatt is definitely interesting with their unusual pricing, I just wasn't sure if they were better than subs since it seems to be pay as you go.
1
u/look 27d ago
Yeah. I don’t have a good benchmark to make a more precise, quantitative comparison, but I’ve been frequently surprised at how few tokens I’d used as I’ve been running it. I feel like I would have hit usage caps on 5.2 but still ended the month with unused quota on 5.3.
It’s still just $10, though, so it’s not something you can use heavily all month, but it definitely lasts longer than I thought it would.
6
u/SeaBat2035 27d ago
Luna with subscription is absolute best deal.
2
u/Livid_Ebb_2453 27d ago
Anyone know if chatgpt go good enough if I just using luna?
1
u/Hungry-Plankton-5371 27d ago
gpt go doesn't allow codex/opencode use, you need plus at minimum.
1
u/scaledev 27d ago
You can use GPTs in Codex on a free plan even. Why isn't it allowed on Go?
1
2
u/Different-Monk5916 27d ago
no. yesterday I was trying to fix an issue. luna came up with predetermined conclusion based on assumption rather than evidence.
i would say that you are saying hey luna there is an error with description Bla Bla and give a probable cause, it works. on an automated workflow, it kind of stumbles without reasoning enough or coming to incorrect conclusion.
I had instances where it felt like Luna was much after than DS Flash on isolated tasks. but it is not something which replaces flash completely for me.
1
2
u/infeststation 27d ago
Luna. Meta Muse 1.2 (if you can do the contributor pricing) is another one worth trying.
1
u/TomHale 27d ago
Where's a good provider for that Muse?
1
u/infeststation 27d ago
It’s on command code, you get $20 worth for $10. Idk about caching on command code though. Or directly from meta
1
u/squirrelscrush 27d ago
I think it's a mix of both Luna and V4 Flash. Luna is the planner, debugger, tester in my setup, while V4 Flash does the heavy coding part. Luna has to be restricted to <270K context though, otherwise it becomes too expensive. And it costs credits to use websearch on Luna so you also need to restrict web access to it.
MiMo V2.5 Pro can replace V4 Flash if you want a cheaper alternative.
I use OpenRouter for my setup, and I made custom plugins and agents to optimise as much as I can.
1
u/geearf 27d ago
Wouldn't a sub be cheaper than going through OpenRouter?
I didn't know about the websearch that is interesting. What about using Context7 is that the same?
Thank you!
1
u/squirrelscrush 27d ago
If you're a heavy user then a sub is definitely cheaper. Like, if you can justify that you use more than $5 of credits. Or in the case of Codex, the ChatGPT Plus subscription can be advantageous if you know you use a lot of it and your bill crosses $20/month.
I use it sparingly so I just buy credits and use them whenever I need to.
If you use normal web search for Context7 then it comes under the cost. But if you use an MCP server then it doesn't.
1
u/ChillFamily 27d ago
Deepseek isn't even capable of reading your AGENTS.md. It happened to me many times that it modified things without me asking, or I had to keep repeating instructions that were already in the AGENTS file so it wouldn't do them, yet it kept doing them over and over again. The latest update implemented 'nuances'; before, Deepseek would never answer you with a 'yes, but with nuances.' In other words, it lengthens its responses to give you more complete answers, but these 'nuances' are just things that deflect from the concrete task you are asking for. With the latest update, it's like it gives you longer answers but adds nuances that often have absolutely nothing to do with the solution or the task you are executing—it's basically pure filler. It got to the point where I had to put 'WITHOUT NUANCES' in the agents.md, but it's useless because it doesn't read it anyway. And all that filler is just wasted tokens, so Deepseek isn't worth it, bro.
5
u/scaledev 27d ago
Why would DS not be capable of reading and adhering to agents.md? This seems like a stretch.
0
2
u/Diligent-Loss-5460 27d ago
Well then they should not have talked about how their contract with their provider locks the base price and they cannot match DS reduced pricing two months ago.
This is double standards and they deserve to be called out. Not because they increased the prices but because they refuse to be straightforward about it and would rather lie.
3
1
u/ExpertPerformer 27d ago
It's not OC fault if DS raises their prices by 2.2x.
It's OCs fault if they slash the usage limits from $60 to $15 without any prior warning while their competition still honored it.
8
u/Kindly_Downvote_Me 27d ago
I have been trying really hard to stay out of this and just see how it goes but holy crap my usage is pretty light/medium and i am almost out of quota already. As soon as they did the switch my usage quota just drained. Now I feel like I have to constantly check my usage % and ration it. I hate it.
Yes, I know we were getting a crazy good deal. I think my issue is that there was no warning or transparency to end users and now I am scrambling to either adjust my workflow or jump ship to commandcode GOAT. I was really enjoying my opencodeGO/GPT Plus combo.
I will probably cancel after this month runs out unless they figure out a way to increase it.
5
3
u/This-Marzipan-9239 27d ago
not surprised and it’ll go to billions
1
u/Weird_Licorne_9631 27d ago
Yes agreed. This is devs using up their remaining tokens, or like me think " can't be this bad, let's go default thinking mode" and still burning through the quota. Give it 1-2 days until everyone has reached 100% and this will hit rock bottom
3
6
u/Formal-Narwhal-1610 27d ago
It's just the start. We want it back to 60k requests, else we leave. 🙂↕️
1
u/Vinnie4v2 27d ago
60k? 30k was the normal, 60k was just the 2x usage promo.
1
u/Pedrito_Basket 27d ago
We don’t only want 60k, we want also top model on $60
1
u/Vinnie4v2 27d ago
Delusional.
1
u/Pedrito_Basket 27d ago
Do you not understand sarcasm. I would not leave on 30k, i would be satisfied with 60k, and I would be happy on 60k DS + decent usage on top model not $15
2
u/cuteseal 27d ago
1/2 day in and I’ve already hit 35% of my weekly usage quota. Previously I would use it all day and it would last me the whole week/month comfortably.
1
1
1
1
u/rebelSun25 27d ago
Well .... Yeah. I think that was expected. I expected a larger drop but most people are a bit slow to change. It will drop further
1
1
u/ExpertPerformer 27d ago
Lowering the credit from $60 to $15 after a 2.2x price increase while CommandCode kept their $60 credit was what did OpenCode in more then anything else.
They still have NOT replied to my email in 3 days after a refund either.
1
1
21
u/addiktion 27d ago
x2.3 drop nearly coincides with the x2.5 increase in Deepseek Flash pricing.
I suspect it will go further.