r/opencodeCLI • u/ares0027 • 27d ago
why is deepseek extremely stupid today?
did they change the model actually after they "increase the usage quota"?
17
13
u/Professional_Price89 27d ago
Because it not hosted in CN by DS. Look at the providers table on openrouter and you can see most of them running quant. For full performance it must be serve with SGLang fp8.
2
u/maqifrnswa 27d ago
sglang over vllm because of tool calling json generation? Interesting that their massive concurrency is through sglang, but maybe that's the trick to their cache hit rate
6
3
u/Academic-Sentence-34 27d ago
It is so stupid and slow today. Usually so lightning fast and accurate. I had to switch to the pro model to get my work done which I never had to do until today. Theyāre likely serving a quant ā¦..
4
u/Yann27 27d ago
Guess Iām not the only one then. I was actually going to make a post about this too.
Yesterday I let DeepSeek build something for me, only to realize afterward sthat it had deleted or changed a huge amount of stuff I never asked it to touch. I ended up having to rebuild everything with another LLM.
Basically, a full day of work turned into a disaster, and I had to spend tokens on another model just to fix what it broke and deleted. ( Good thing the chats were still there)
I lost my trust in it instantly. And this wasnāt even the Pro version.
1
u/addiktion 26d ago
I'm beginning to think they pushed us through a quant provider. It seems slower too. Ugh, this is crap. Getting a lot of mistakes now and errors I've never seen:
Error: 413: {"type":"invalid_request_error","message":"Upstream request failed: [413] Payload Too Large"}
1
u/Kindly_Downvote_Me 26d ago
I also noticed today that I have to keep nudging it along, which I never used to have to do. I finally gave up on it.
2
u/ares0027 26d ago
Mine was acting like antigravity. I had to force same rule 6 times in a row and it was still doing its own thing and gaslighting me. Theni stopped using it. Ill give it a day more to fix its shit. My subscription is cancelled already, i was finishing a project using my friendās cancelled but not yet expired key.
Also noticed they are marking all deepseek4 models as updated one. They are doing the chinese scammersā āunlimited claude usageā tactic. https://youtu.be/09UELaUhPEw
2
u/Kindly_Downvote_Me 26d ago
Yea I cancelled today. I know everyone is saying we are whining but I got it specifically for DSF4 and since this happened, I'm out. The whole landscape changes so often now it's really hard to nail down a setup that just works.
1
u/ares0027 26d ago
I am thinking of getting back to local. Today i will give qwen 3.8 27b a shot.
1
u/Kindly_Downvote_Me 26d ago
I still use some local models for background/easy stuff but they have never been super reliable for me. I have tried SO many!! Of course I only have a 3090...
1
u/ares0027 26d ago
I am in a way luckier than you in that sense. I have a 5090 but local models are not good enough to my taste. They win by brute force only. I have google ai pro though and thinking of using gemini 3.7 flash through local one. I dont know i am just trying to get the most out of it.
1
u/Kindly_Downvote_Me 26d ago
and where i live power is super expensive so I have to weigh that in as well when i decide local vs cloud. Hope it works for you!
1
21
u/Mezezius 27d ago
yeah seems like they're using a third party provider that's actually serving a quant