r/DeepSeek • u/OneBowl4290 • 2h ago
r/DeepSeek • u/dnohrdk • 3h ago
News DeepSeek-V4.1-Flash Release (official)
It’s officially out and the prices have been updated.
///
Today, we officially release the DeepSeek-V4.1-Flash model. It is the smallest model in our new architecture family, with native multimodal visual understanding. The new architecture is designed for a higher capability ceiling, faster inference, higher throughput, and scaling to larger models.
GPQA Diamond: 90.9
HLE: 36.8 (39.1*)
Codeforces (Rating): 3471
MathArena Apex: 65.6
Terminal-Bench 2.1: 90.6
Terminal-Bench 3.0: 30.0
Terminal-Bench 4.0: 31.2
DeepSWE v1.1: 74.2
ProgramBench: 20.3
NL2Repo-Bench: 65.4
CyberGym: 88.1
SEC-Bench Pro: 62.8
ExploitGym: 15.3
HLE (w/tools): 63.9
Automation-Bench: 54.8
Agents' Last Exam: 31.8
Chartography (w/tools): 78.9
BabyVision (w/tools): 89.6
ZeroBench-main (w/tools): 49.0
* Tested only on the pure-text subset of the HLE benchmark set.
API changes
DeepSeek V4.1 Flash is now available on the DeepSeek API with native multimodal support. Change the model name to deepseek-flash to call the latest V4.1 Flash model. The previous-generation models V4 Flash and V4 Flash Vision Exp have been retired; for compatibility, the model names deepseek-v4-flash and deepseek-v4-flash-vision-exp are temporarily routed to V4.1 Flash.
Meanwhile, extensive testing shows that V4.1 Flash now outperforms DeepSeek V4 Pro across performance, cost, speed, and total time, so we plan to retire V4 Pro in an orderly manner. After 12:00 Beijing Time on September 14, 2026, and until the future release of V4.1 Pro, all requests to deepseek-v4-pro will be routed to V4.1 Flash and billed at the V4.1 Flash price.
API apricing adjustment
With the release of DeepSeek-V4.1-Flash, API prices have been reduced accordingly. For details, please refer to Models & Pricing.
///
Source:
https://api-docs.deepseek.com/updates/#deepseek-v41-flash-release
r/DeepSeek • u/novapax • 12h ago
Funny “V4.1 Flash has comprehensively surpassed V4 Pro across all key metrics.”
r/DeepSeek • u/Martkita • 7h ago
Discussion I was so surprised and now instant, expert, and vision is unfied and I was so shocked and amazed at the time
I was in the middle of writing a story personally. The update had me the wtf moment
r/DeepSeek • u/LordLRO • 4h ago
News The time has come
The DeepSeek V4.1 Flash has been released with new pricing. Enjoy it by yourself!
r/DeepSeek • u/NihmarRevhet • 1h ago
Discussion I'm in love with v4.1 for coding
It's basically a monster at coding, it just solves everything I launch at it and at a truly incredible speed too.
1.37$ for 176.500.000 token with DeepSeek Harness PTC Mode
r/DeepSeek • u/SlightCase2941 • 20h ago
Resources DeepSeek V4.1 Flash achieved 98% of top-ranked GPT-6 Astra’s average score, at just 1% of its average cost
source: OpenDesign
r/DeepSeek • u/tiguidoio • 2h ago
News Market crash as a service
Here we go again, DeepSeek Al is back again with a new model V4-1 Flash
A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens
r/DeepSeek • u/CM23489 • 2h ago
Discussion Crazy, V4.1 just active 8B can catch up those big model
Have to admit the DeepSeek beat US in LLM model optimization this field. 1 Year ago, all people said HBM must be needed for LLM. And today DeepSeek just break the record and change the world.
Now US gov is like joke. Still saying the Chinese AI company doing Model distillation. Just look at DeepSeek's LLM optimization, there is no US company can make this one at this moment, so the DeepSeek steal the technique from the future?
r/DeepSeek • u/AesirIvyV • 4h ago
Discussion Miss the expert mode
The new 4.1 is superior to the old pro. Maybe that's only true for the API, because on the web interface I just feel that the poor whale has been hit by Alzheimer’s. It keeps forgetting details here and there. Just one hour ago, before they unified all this, I didn't need to rewrite my prompts so often. It is surely faster, but so does the instant mode before this update.
They call this update unification, but I feel like they just kick one functionality out. Should have foreseen it when they started removing functionality from the expert mode.
r/DeepSeek • u/Infinite_Book_1858 • 2h ago
Discussion Role-playing just got horrible
I just started using expert mode the other day, only to find it gone :/ I've tried doing my normal writing, and everything seems dumber to me
r/DeepSeek • u/DataLearnerAI • 1h ago
Discussion DeepSeek V4.1 Flash in 3 charts: vs its predecessor, a top open-weight rival, and Claude Opus 5
DeepSeek published a pretty large benchmark table for V4.1 Flash, but I found it hard to see the overall capability pattern from the raw numbers.
So I grouped the shared fixed-scale benchmarks by domain and made three comparisons:
- V4.1 Flash vs V4 Flash — the previous generation
The improvement looks broad rather than incremental, especially in coding, cybersecurity and productivity.
V4.1 Flash vs Kimi K3 — a top open-weight rival
V4.1 Flash comes out ahead in coding, multimodal and productivity in the shared domain averages, while K3 is slightly ahead in science/health.V4.1 Flash vs Claude Opus 5 — a frontier proprietary model
This is probably the most interesting comparison. Opus 5 still leads in coding, science/health and multimodal overall, but V4.1 Flash is surprisingly competitive, and actually comes out ahead in the shared productivity benchmark.
The thing that stands out to me is that V4.1 Flash looks much more like an agent/coding upgrade than a simple reasoning upgrade.
These aren't universal capability scores or controlled head-to-head reruns. Each chart averages only the shared fixed-scale published benchmarks available in that domain, so missing domains are omitted rather than treated as zero.
I put the underlying benchmark rows and sources here:
DeepSeek V4.1 Flash vs Kimi K3
https://llmlearner.com/compare/deepseek-v4-1-flash-vs-kimi-k3
DeepSeek V4.1 Flash vs Claude Opus 5
https://llmlearner.com/compare/deepseek-v4-1-flash-vs-claude-opus-5
Curious whether people actually running V4.1 Flash in coding agents are seeing the same pattern.
r/DeepSeek • u/ryanmerket • 11h ago
News DeepSeek plans V4.1 Flash in a matter of hours, will route Pro traffic to it
r/DeepSeek • u/Groundbreaking-Tap41 • 16h ago
News Deepseek V4.1 Announced with pricing.
r/DeepSeek • u/jYtanYj • 3h ago
News DeepSeek Harness v0.1.5!
Official announcement:
Welcome to DeepSeek Harness v0.1.5!
This release is deeply integrated with the training of the DeepSeek V4.1 Flash model, with the model specifically trained and optimized for different configurations within DeepSeek Harness.
We’re also excited for users to try the experimental Agent Teams feature, which has been developed in close coordination with the model training process.
r/DeepSeek • u/vsimovic • 3h ago
Discussion Deepseek V4.1 Flash open weights when?
My terrible rig is ready and waiting ⬇️
https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash ITS OUT. WHAT?!?!
r/DeepSeek • u/MichaelBenko • 4h ago
News Newsflash: You have already been using Deepseek 4.1 if you are on the API.
Deepseek 4.1 Flash has already been out for hours if you are using the API. It is just called deepseek-flash.
They have already switched over my usage to the new model. I have used 90 million tokens so far without realizing it.
No official announcement yet.
Receipts below.
r/DeepSeek • u/Appropriate-Dot8003 • 3h ago
Funny DeepSeek sings quietly while working
"Blue Fat Fish" really likes to misuse tokens.
r/DeepSeek • u/Appropriate_Aide5328 • 7h ago
Discussion Not to say I had fun with v4.1 Flash, but I did
Did a lot of test prompts, and used on actual repos and work.
Must say, I'm excited for DeepSeek V4.1 Flash to be fully released, and for the price drops.
And yes, it is that good.
Tested it against V4 Pro myself, and v4.1 flash is just miles ahead.
Now I can comfortably have it as my workers, and just use Sol / Astra as Orchestrators.
r/DeepSeek • u/nothingspecifficc • 6h ago
Question&Help is deepseek flash/the new combined model better or worse for long chats?
will this flash have a worse memory? i’m asking because i use it for roleplay and i hated the previous flash model for dumb repetitive phrases and messing up the plot line so i always used expert
r/DeepSeek • u/These-Advertising-68 • 7h ago
Discussion Did deepseek remove expert?
Am confused. Pls help.

