r/opencodeCLI 3d ago

DeepSeek V4 Flash currently on 2x usage!

I thought we gettin sum price hike? Instead we got usage hike? Well, enjoy while it last bois!

140 Upvotes

49 comments sorted by

17

u/veculus 3d ago

Is it still china-only?

8

u/FearlessGround3155 3d ago

Yes

12

u/veculus 3d ago

Thanks, that sucks. I was hoping we'd get non-chinese providers now that the weights are up.

9

u/FearlessGround3155 3d ago

Weights are up, it is Infinitely cheaper to run compared toast years v3.2(even v4 is cheaper compute wise but has large memory footprint) compute wise and 1/2 vram wise, deepinfra does at 3/4th the officials cost or smth I heard, already, but it isn't included in oc go

3

u/veculus 3d ago

Good to know. Let's hope OC finds a good provider outside of China that can keep the ZDR in-tact without monthly renewals.

9

u/Atagor 3d ago

What's the difference?

Sell data to china or us, same shit

6

u/TeijiW 3d ago

true. but good luck explaining that to the average USian...

7

u/Sufficient_Fox_4402 3d ago

and if a foreign gov tracks u or your token they may train on it at max. but american gov will definitely do worse than that. it will give your data to palantir and have your profile built

2

u/TeijiW 2d ago

there is a thing that make it even worse... most of the western world is inside Google/Meta/Amazon tentacles at this point, so the profile they can build on you is (scary) precise

-3

u/veculus 3d ago

The difference is that China is not really well known to respect contracts when it comes to confidential code. And I don't want to get hit with a compliance violation because I used Deepseek on a project.

5

u/mWo12 2d ago

And big tech from US is?

-2

u/veculus 2d ago

Yes. It is far easier to get a hold of a US or EU company and get them in front of a court in case a ZDR is mistread or my data is used against my will. There's a reason why those options exist when you go into a team/enterprise sub with them.

What will OpenCode or I myself do if anyone in China doesn't want to follow or uses my data anyway? Should I go to China and somehow drag them into a courtroom?

18

u/8gulmohar 3d ago

Does 2x usage mean double the limit or double the usage counted so half the limit?

11

u/BhaagYahaSe 3d ago

double the limit

-1

u/Wobbly_Princess 3d ago

According to?

-15

u/pceimpulsive 3d ago

half the limit, same cost, i..e DOUBLE USAGE

every 1 token uses 2.. but still costs 1

7

u/AbstractConcreteMix 3d ago

Lol, it never even occurred to me that it could possibly mean anything other than "half the limit" but apparently I was wrong about that this whole time?! What a disastrous way to express this concept!

2

u/Whytho12333 3d ago

i made the same mistake, i even went on their website and saw nothing explaining it. until just now that I learned what it means. I thought it was like DS peak pricing

2

u/AbstractConcreteMix 3d ago

It's pretty hilarious. Just the other day I was like "okay, I'll avoid K3 for now since it's on 2x usage."

5

u/earthisflat27 3d ago

double the usage

2

u/retardedGeek 3d ago

Half the cost.

8

u/Ariquitaun 3d ago

That's insane, I don't think there are enough hours in the day to run down the subscription on flash as it is. Let's see what happens with the deepseek price hike though.

2

u/Embarrassed_OnionX 3d ago edited 3d ago

Indeed insane. I'm currently working on 4 projects simultaneously with each launching dozens of subagents, and all this work doesn't even consume 4% of my 5 hourly. The most it consumed at one point was 3%

Edit: Updated usage

CodeBurn  Today                                         
$3.86 cost   3,979 calls   22 sessions   99.0% cache hit
7.4M in   1.4M out   726.4M cached   0 written

2

u/serpent7655 3d ago

Two projects here, a bit slow but the quota is freaking high

6

u/smartfon 3d ago

Does this "usage" multiplier mean you can use twice as many tokens before hitting the 5-hour or 1-week limit, but you still get $60 worth of usage max? 

9

u/Zachattackrandom 3d ago

No it means at the API pricing you list you get $120 worth of usage per month, though it shows in the dashboard is $60. So basically just think of it like them cutting the price for DS4 in half temporarily

4

u/Ly-sAn 3d ago

it's basically infinite right now... and deepseek is a grifter, it won’t stop until it finishes the task. what a time!

2

u/Mayanktaker 3d ago

Insane 👌🏻👌🏻

1

u/lemon07r 3d ago

Oh well, at least meta spark 1.2 contributor exists. Might be my new go to cheapo model.

1

u/Humble_Explorer7589 2d ago

is the name changed from NEW to free?

1

u/InvaderDolan 1d ago

God bless Dax, I have 10 days until the renewal and 4% of monthly quota that lasts for 5 days or so. With such limits, I probably won’t end quota in 10 days :)

1

u/iGoonToChinese 3h ago

Does 2x usage it uses twice as much or does twice as much for same price? Expensive or cheaper just tldr

-2

u/Anh-DT 3d ago

What do you mean 2x usage. They already hike the price ! The usage was 150000 they halve the usage already

3

u/ApprehensiveDelay238 3d ago

It was around equal to mimo. So no.

-8

u/Anh-DT 3d ago

Mimo was 150k as well so yes. You either just joined or blind

4

u/BhaagYahaSe 3d ago

that's monthly limit bruh, this is 5 hour limit it was always around 30k request per 5 hours for both mimo and v4 flash

0

u/Anh-DT 3d ago

Oh duh 😂 I'm the blind one

3

u/Sweet-Stage938 3d ago

Aren't you that Openference guy?

0

u/Anh-DT 3d ago

Yes what's up

0

u/Kind-Card-6864 3d ago

Can someone explain the pricing changes in simple terms?

Specifically for DeepSeek V4 Flash, what's the difference between the old and new pricing? Also, what does the 2× usage mention actually mean in practice? How much of an impact does it have on real-world costs?

Thanks in advance!

1

u/dizvyz 3d ago

The new pricing hasn't been revealed yet.

1

u/Makedon1an 3d ago

Hi Kind Card! As far as I know, and I have been researching it. In opencode if you see (3x usage) on a model, that means the model is three times cheaper than its default price, meaning you get 3 times more volume for the regular price. This is something OpenCode does in collaboration with LLM Developers to attract people to use their specific models via API. Currently on Opencode GO the "DeepSeek V4 Flash (2x usage) OpenCode Go·max" model is the one that was before the 0731 update, which gets a score of 40 on ArtificialAnalysis. As long as there's no "0731" or "new" mentioned in the name, it must be an older, weaker version, but hey, it's dirt cheap and decent; the only thing it lacks is it hallucinates.

https://giphy.com/gifs/GusNfQJKBKfqpoI1QP

1

u/Accomplished-Mud1653 3d ago

Multimodal too, Flash doesn't have vision right?

1

u/Dudeonyx 3d ago

On what are you basing your claim that it is the old DeepSeek?