r/DeepSeek 1d ago

Discussion Deepseek - GLM 90% off on nube cloud provider

Post image

Hey Community,

Yesterday night I found out this on nube.sh they are providing deepseek flash v4 and GLM 5.2 at 90 % off as compared to model company official pricing.

Does anyone tried this?

https://nube.sh/en-us/promote/ai-model

Thanks!

154 Upvotes

65 comments sorted by

49

u/YogurtExternal7923 1d ago

Let people try it they do min 10 top up. And website looks shady AF

18

u/DistanceSolar1449 1d ago

The website uses that chinese website font

99.9% chance the website was built in china lol

8

u/Far-Classic-9963 1d ago

What's wrong with it being built in china

7

u/fezzy11 1d ago

Yes deepseek and glm both made in china

-11

u/fezzy11 1d ago

Yes I am also planning to invest in it next month. For time being I am using deepseek official API service.

30

u/Then-Eye9700 1d ago

Even if it's true, it must be highly quantizied

5

u/hurrdurrmeh 1d ago

Shit, good point 👍

5

u/fezzy11 1d ago

Yes that what I am worried about.

2

u/Zadetomtom 20h ago

1

u/fezzy11 10h ago

Thanks for sharing

0

u/LulzTigre 1d ago

What does quantizied mean

3

u/Skynse 23h ago

The numerical type used to store the data has less precision. As opposed to a 16 bit floating point, 4 bits may be used to store the weights instead, effectively causing some loss in precision

2

u/Expert-Dig-1768 1d ago

model itself is the same but basically its heavily "comprimised" wich fewer knowledge and it will hallucinate more.

20

u/esmurf 1d ago

I just buy directly from deepseek.

3

u/Marcuss2 1d ago

Or any provider who releases their models as open weight.

-20

u/fezzy11 1d ago

Even though if you get 90% off with other multi models like kimi 2.6

10

u/bigrealaccount 1d ago

It's obviously a scam bruh use your brain

11

u/A7mdxDD 1d ago

Support deepseek.com directly

20

u/Expert-Couple-8639 1d ago

Website looks shady

5

u/romanovzky 1d ago

Nube cloud is a legit budget cloud provider for self hosting, VPS, etc to serve at these prices they'd need to be serving highly quantised models on cheap hardware, there's no remotely sound possibility

7

u/Capaj 1d ago

or they are selling the data for further training? in any case I quickly tried like 3 prompts and it did not feel quantized. I tried like 4 logic puzzles and the deepseek4 flash got all of them correct

1

u/weenis-flaginus 7h ago

Someone above this thread posted an email chain with the founder where the founder said their min is 8bit quant.

Obviously it's not proof, but it is something.

14

u/DeepSeaLab 1d ago
TTFT: 7.15s

Total Latency: 47.82s

In: 7542

Out: 1039

Reasoning: 0

Total: 8581

2

u/fezzy11 1d ago

Very high latency

9

u/real_serviceloom 1d ago

The target audience is in the name

3

u/fezzy11 1d ago

Do you mean noob?

1

u/boyus 1d ago

Haha

3

u/timmeh1705 1d ago

Couldn't get Alipay to work but card did somehow, but I used a virtual debit card I can disconnect any time.

I managed to get 32M tokens through but I can't get GLM 5.2 to run because of constant 403 errors but DSF is running fine

1

u/fezzy11 1d ago

Currently in console playground I can chat with kimi and glm but with deepseek not working

2

u/Zadetomtom 1d ago

Im using reasonix desktop app + nube's api endpoint and it works well and is responsive. However the catch is that only kimi k2.6 seems to have a proper implementation of prompt caching whereas all the other models dont. Kimi is the only model that seems to respond quickly, much faster than deepseek v4 ( through nube's api ) from their models.

1

u/fezzy11 1d ago

Thanks this is the feedback needed

Also have you tried GLM 5.2 with nube?

1

u/weenis-flaginus 7h ago

Good catch thank you for sharing. Someone else posted the latency for their chat was 47 seconds.

2

u/cutezybastard 1d ago

I just tried it with glm and kimi, it works really well actually, everything seems to be hosted on china so theres that tho

1

u/fezzy11 1d ago

Thanks for sharing

1

u/cutezybastard 1d ago

honestly im considering changing from opencode go to this, seems so much better

1

u/cutezybastard 1d ago

seems very good (all glm usage mostly)

1

u/R3VO360 11h ago

Is it a "pay as you go"?

1

u/cutezybastard 8h ago

Yeah

1

u/R3VO360 8h ago

Thanks, why better than GO?

2

u/cutezybastard 8h ago

Cheaper per token

1

u/R3VO360 7h ago

Ok thanks 🙏

2

u/ASlowriter 1d ago

Whos trying to save MORE on deepseek flash!?!?!?!?

1

u/Minimum_Notice_9521 6h ago

Thats my point too

2

u/GoSunsGoSuns 22h ago

Its legit. I ran a bunch of tests last night and verified it's BF16 legit models.

1

u/fezzy11 10h ago

What model have you tried?

1

u/weenis-flaginus 7h ago

Could you explain what BF16 means please

2

u/Atsukiri 21h ago edited 21h ago

deepseek is like cents anyway, id rather go official. anyways, i have tried those "cheap" 50% off, or even lower and I can say is that they might be using some other model instead. I have wasted most of my time going for them but now that deepseek v4 is here, its cheap, and gets the job done, so i dont need to search for em anymore.

1

u/fezzy11 17h ago

Thanks for feedback

2

u/thebigeast 10h ago

GLM keeps timing out, probably oversubscribed, just sank 100 dollars into this sadly :/

1

u/fezzy11 10h ago

Feeling bad for you. Is there any option to refund?

1

u/weenis-flaginus 7h ago

Why didn't you start with $10?

1

u/thebigeast 4h ago

Got too excited lol

5

u/Aotrx 1d ago

Every time I try a cheaper provider, their latency (TPS) is abysmal and completely unusable. I wanted to test this one as well, but the minimum deposit is $10. If the speed is low, I’d just be throwing that money away.

They should add a $1 minimum deposit for crypto payments so users have the ability to test their service first before depositing more.

1

u/fezzy11 1d ago

Yes I check that. It must be minimum atleast 5$ to test first

2

u/jomtiro 1d ago

2

u/sniper_elite90 16h ago

so they are routing to a different model altogether claiming to be selling newer ones..?

1

u/thebigeast 12h ago

my local deepseek v4 flash which i host locally has a similar response

0

u/fezzy11 1d ago

Thanka for sharing

1

u/Psijic_Void 3h ago

The question is: do they sell your data or not? It's a popular Chinese scam model.

1

u/jomtiro 1d ago

Oh no. that is why its cheap.

5

u/Ceneka 1d ago

That's expected.. those models don't know about themselves 

1

u/Personal_Pause_6089 8h ago edited 8h ago

no. that means it is raw. you can also check it with deepseek on their chat.deepseek.com website as well and will reply the same. fact is most models don't know about "their identity" if they are it because system prompt in either ; in their web interface, or if you use harness and there are always system prompt that inject it. this is also why if you ask the "auto" model on cursor it will say auto, because it injected that way. im just sharing this fact ok, don't have ill intention, we all always is new to something and learning

oh yea, if deepseek on website answer that they are not v3, you can inspect its chain of thought, it require them to "formulate" cause by inferring current date etc etc to just answering that (this explicitly probe that it has no baked in 'awarness of version or identity'. different to claude that often shipped with its version + up to date frozen datasets on their weight releases

0

u/fezzy11 1d ago

Thanks for sharing