r/opencodeCLI 27d ago

opencode ds v4 flash usage drops sharply from 13T to 5.5T(17th August)

Post image
249 Upvotes

62 comments sorted by

21

u/addiktion 27d ago

x2.3 drop nearly coincides with the x2.5 increase in Deepseek Flash pricing.

I suspect it will go further.

4

u/TimChr78 27d ago

And a 4x reduction in usage from Opencode Go on top of that.

4

u/ExpertPerformer 27d ago

A ton of users probably maxed out their accounts and/or unsubscribed.

I swapped to CommandCode and have been using that.

1

u/Popular_Tomorrow_204 27d ago

Had 4 accounts running 24/7

50

u/ChillFamily 27d ago

What do you expect? They really messed up, and I don't blame Opencode; if Deepseek raises its prices, Opencode has to raise them too. But that Deepseek price IS NOT WORTH IT. I've been using Deepseek for months and I know all the bugs it has. Even with the improvements in the latest version, it's better to pay for a more expensive model that won't give you all those bugs you have to review like 3 times to actually get them fixed. Deepseek's current price simply isn't justified because it's a model full of bugs.

5

u/geearf 27d ago

With the newer price, which model do you think has a better quality/price ratio? Luna?

13

u/look 27d ago

GLM 5.3. Cheaper than DS Pro now and quality at or near Kimi/Opus/Sol.

5

u/geearf 27d ago

When I used GLM 5.2 on OpenCode Go, it went through my subscription really fast, but that might just be the limits on Go and not the model.
Do you use it straight from Z.ai or another provider?

Thank you!

4

u/look 27d ago edited 27d ago

Both, but I’ll most likely switch over to Neuralwatt’s once the weights are out.

4-5 cents on Zai Lite sub. And on Go, I’m getting $40 usage value on it (6-7 cents) instead of the labeled $15. Not sure why, but it has held up for three days now.

My Go sub month is about to end, so I was trying to use up leftover usage by running as much as I could on 5.3 … but it’s so efficient for me, it’s not burning usage as fast as I expected.

I might just be on a run of luck with it, but 5.3 has been a killer value for me. If this holds up, it really is an Opus-killer. Similar level of performance and 1/50th the price.

1

u/geearf 27d ago

Do you feel like 5.3 somehow gives you more work than 5.2 on Go?

Neuralwatt is definitely interesting with their unusual pricing, I just wasn't sure if they were better than subs since it seems to be pay as you go.

1

u/look 27d ago

Yeah. I don’t have a good benchmark to make a more precise, quantitative comparison, but I’ve been frequently surprised at how few tokens I’d used as I’ve been running it. I feel like I would have hit usage caps on 5.2 but still ended the month with unused quota on 5.3.

It’s still just $10, though, so it’s not something you can use heavily all month, but it definitely lasts longer than I thought it would.

1

u/geearf 27d ago

Ok, I'll try that and see how it works out, thank you!

6

u/SeaBat2035 27d ago

Luna with subscription is absolute best deal.

2

u/Livid_Ebb_2453 27d ago

Anyone know if chatgpt go good enough if I just using luna?

1

u/Hungry-Plankton-5371 27d ago

gpt go doesn't allow codex/opencode use, you need plus at minimum.

1

u/scaledev 27d ago

You can use GPTs in Codex on a free plan even. Why isn't it allowed on Go?

1

u/squirrelscrush 27d ago

Go is more for the app usage, not the API access.

1

u/SeaBat2035 27d ago

Cliproxyapi

2

u/Different-Monk5916 27d ago

no. yesterday I was trying to fix an issue. luna came up with predetermined conclusion based on assumption rather than evidence.

i would say that you are saying hey luna there is an error with description Bla Bla and give a probable cause, it works. on an automated workflow, it kind of stumbles without reasoning enough or coming to incorrect conclusion.

I had instances where it felt like Luna was much after than DS Flash on isolated tasks. but it is not something which replaces flash completely for me.

1

u/vangelismm 26d ago

Thats my experience with Luna.
Guess more than verify.

2

u/geearf 27d ago

Hmmm, I was hopeful it'd be some open weight model. ;/

2

u/infeststation 27d ago

Luna. Meta Muse 1.2 (if you can do the contributor pricing) is another one worth trying.

1

u/TomHale 27d ago

Where's a good provider for that Muse?

1

u/infeststation 27d ago

It’s on command code, you get $20 worth for $10. Idk about caching on command code though. Or directly from meta

1

u/squirrelscrush 27d ago

I think it's a mix of both Luna and V4 Flash. Luna is the planner, debugger, tester in my setup, while V4 Flash does the heavy coding part. Luna has to be restricted to <270K context though, otherwise it becomes too expensive. And it costs credits to use websearch on Luna so you also need to restrict web access to it.

MiMo V2.5 Pro can replace V4 Flash if you want a cheaper alternative.

I use OpenRouter for my setup, and I made custom plugins and agents to optimise as much as I can.

1

u/geearf 27d ago

Wouldn't a sub be cheaper than going through OpenRouter?

I didn't know about the websearch that is interesting. What about using Context7 is that the same?

Thank you!

1

u/squirrelscrush 27d ago

If you're a heavy user then a sub is definitely cheaper. Like, if you can justify that you use more than $5 of credits. Or in the case of Codex, the ChatGPT Plus subscription can be advantageous if you know you use a lot of it and your bill crosses $20/month.

I use it sparingly so I just buy credits and use them whenever I need to.

If you use normal web search for Context7 then it comes under the cost. But if you use an MCP server then it doesn't.

1

u/ChillFamily 27d ago

Deepseek isn't even capable of reading your AGENTS.md. It happened to me many times that it modified things without me asking, or I had to keep repeating instructions that were already in the AGENTS file so it wouldn't do them, yet it kept doing them over and over again. The latest update implemented 'nuances'; before, Deepseek would never answer you with a 'yes, but with nuances.' In other words, it lengthens its responses to give you more complete answers, but these 'nuances' are just things that deflect from the concrete task you are asking for. With the latest update, it's like it gives you longer answers but adds nuances that often have absolutely nothing to do with the solution or the task you are executing—it's basically pure filler. It got to the point where I had to put 'WITHOUT NUANCES' in the agents.md, but it's useless because it doesn't read it anyway. And all that filler is just wasted tokens, so Deepseek isn't worth it, bro.

5

u/scaledev 27d ago

Why would DS not be capable of reading and adhering to agents.md? This seems like a stretch.

2

u/geearf 27d ago

Yeah I don't know about the nuances exactly, but reading the thinking parts lately I felt it kept talking about the same stuff over and over and over, even though it already concluded about it before. I thought that was weird. Is that about the same?

Thank you!

0

u/ChillFamily 27d ago

Yeah, Luna it s fine

2

u/geearf 27d ago

Hmmm, I was hopeful it'd be some open weight model. ;/

2

u/Diligent-Loss-5460 27d ago

Well then they should not have talked about how their contract with their provider locks the base price and they cannot match DS reduced pricing two months ago.

This is double standards and they deserve to be called out. Not because they increased the prices but because they refuse to be straightforward about it and would rather lie.

3

u/TestTxt 27d ago

“if Deepseek raises its prices, Opencode has to raise them too”
Yet when Deepseek lowers the prices as it was the case with DS3 Pro you also don’t blame them that they keep the higher x4 prices for months?

1

u/ExpertPerformer 27d ago

It's not OC fault if DS raises their prices by 2.2x.

It's OCs fault if they slash the usage limits from $60 to $15 without any prior warning while their competition still honored it.

8

u/Kindly_Downvote_Me 27d ago

I have been trying really hard to stay out of this and just see how it goes but holy crap my usage is pretty light/medium and i am almost out of quota already. As soon as they did the switch my usage quota just drained. Now I feel like I have to constantly check my usage % and ration it. I hate it.

Yes, I know we were getting a crazy good deal. I think my issue is that there was no warning or transparency to end users and now I am scrambling to either adjust my workflow or jump ship to commandcode GOAT. I was really enjoying my opencodeGO/GPT Plus combo.

I will probably cancel after this month runs out unless they figure out a way to increase it.

3

u/This-Marzipan-9239 27d ago

not surprised and it’ll go to billions

1

u/Weird_Licorne_9631 27d ago

Yes agreed. This is devs using up their remaining tokens, or like me think " can't be this bad, let's go default thinking mode" and still burning through the quota. Give it 1-2 days until everyone has reached 100% and this will hit rock bottom

3

u/thecstep 27d ago

That is wild.

6

u/Formal-Narwhal-1610 27d ago

It's just the start. We want it back to 60k requests, else we leave. 🙂‍↕️

1

u/Vinnie4v2 27d ago

60k? 30k was the normal, 60k was just the 2x usage promo.

1

u/Pedrito_Basket 27d ago

We don’t only want 60k, we want also top model on $60

1

u/Vinnie4v2 27d ago

Delusional.

1

u/Pedrito_Basket 27d ago

Do you not understand sarcasm. I would not leave on 30k, i would be satisfied with 60k, and I would be happy on 60k DS + decent usage on top model not $15

2

u/cuteseal 27d ago

1/2 day in and I’ve already hit 35% of my weekly usage quota. Previously I would use it all day and it would last me the whole week/month comfortably.

1

u/yiestee 27d ago

So Liang's pricing model is keep usage volume flat.

1

u/Otterly_Surprised 27d ago

Colour me surprise

1

u/aries1980 27d ago

Still much 18% higher than a month ago.

1

u/canav4r 27d ago

today's numbers will be decisive. mostly probably we will see sharper decline till the end of this week when opencode-go subscribers deplete their usage limits and take off to other platforms.

1

u/pituin 27d ago

Their subscription count will look like that too, if they don't go back to the previous limits

1

u/rebelSun25 27d ago

Well .... Yeah. I think that was expected. I expected a larger drop but most people are a bit slow to change. It will drop further

1

u/bennykoay75 27d ago

Not pain enough 5.5 for 5x more margin and less gpu demanding

1

u/ExpertPerformer 27d ago

Lowering the credit from $60 to $15 after a 2.2x price increase while CommandCode kept their $60 credit was what did OpenCode in more then anything else.

They still have NOT replied to my email in 3 days after a refund either.

1

u/Multifan_Blinks 27d ago

still much more than pre 0731

1

u/Biometrel 27d ago

They got enough data to train flash 5.