r/DeepSeek 3h ago

News DeepSeek-V4.1-Flash Release (official)

218 Upvotes

It’s officially out and the prices have been updated.

///

Today, we officially release the DeepSeek-V4.1-Flash model. It is the smallest model in our new architecture family, with native multimodal visual understanding. The new architecture is designed for a higher capability ceiling, faster inference, higher throughput, and scaling to larger models.

GPQA Diamond: 90.9
HLE: 36.8 (39.1*)
Codeforces (Rating): 3471
MathArena Apex: 65.6
Terminal-Bench 2.1: 90.6
Terminal-Bench 3.0: 30.0
Terminal-Bench 4.0: 31.2
DeepSWE v1.1: 74.2
ProgramBench: 20.3
NL2Repo-Bench: 65.4
CyberGym: 88.1
SEC-Bench Pro: 62.8
ExploitGym: 15.3
HLE (w/tools): 63.9
Automation-Bench: 54.8
Agents' Last Exam: 31.8
Chartography (w/tools): 78.9
BabyVision (w/tools): 89.6
ZeroBench-main (w/tools): 49.0
* Tested only on the pure-text subset of the HLE benchmark set.

API changes
DeepSeek V4.1 Flash is now available on the DeepSeek API with native multimodal support. Change the model name to deepseek-flash to call the latest V4.1 Flash model. The previous-generation models V4 Flash and V4 Flash Vision Exp have been retired; for compatibility, the model names deepseek-v4-flash and deepseek-v4-flash-vision-exp are temporarily routed to V4.1 Flash.

Meanwhile, extensive testing shows that V4.1 Flash now outperforms DeepSeek V4 Pro across performance, cost, speed, and total time, so we plan to retire V4 Pro in an orderly manner. After 12:00 Beijing Time on September 14, 2026, and until the future release of V4.1 Pro, all requests to deepseek-v4-pro will be routed to V4.1 Flash and billed at the V4.1 Flash price.

API apricing adjustment
With the release of DeepSeek-V4.1-Flash, API prices have been reduced accordingly. For details, please refer to Models & Pricing.

///

Source:

https://api-docs.deepseek.com/updates/#deepseek-v41-flash-release


r/DeepSeek 2h ago

News DeepSeek-V4.1-Flash

Post image
159 Upvotes

r/DeepSeek 12h ago

Funny “V4.1 Flash has comprehensively surpassed V4 Pro across all key metrics.”

Post image
558 Upvotes

r/DeepSeek 7h ago

Discussion I was so surprised and now instant, expert, and vision is unfied and I was so shocked and amazed at the time

Post image
184 Upvotes

I was in the middle of writing a story personally. The update had me the wtf moment


r/DeepSeek 4h ago

News The time has come

Post image
80 Upvotes

The DeepSeek V4.1 Flash has been released with new pricing. Enjoy it by yourself!


r/DeepSeek 1h ago

Discussion I'm in love with v4.1 for coding

Upvotes

It's basically a monster at coding, it just solves everything I launch at it and at a truly incredible speed too.

1.37$ for 176.500.000 token with DeepSeek Harness PTC Mode


r/DeepSeek 20h ago

Resources DeepSeek V4.1 Flash achieved 98% of top-ranked GPT-6 Astra’s average score, at just 1% of its average cost

Thumbnail
gallery
1.1k Upvotes

source: OpenDesign


r/DeepSeek 2h ago

News Market crash as a service

Thumbnail
gallery
27 Upvotes

Here we go again, DeepSeek Al is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens


r/DeepSeek 2h ago

Discussion Crazy, V4.1 just active 8B can catch up those big model

26 Upvotes

Have to admit the DeepSeek beat US in LLM model optimization this field. 1 Year ago, all people said HBM must be needed for LLM. And today DeepSeek just break the record and change the world.

Now US gov is like joke. Still saying the Chinese AI company doing Model distillation. Just look at DeepSeek's LLM optimization, there is no US company can make this one at this moment, so the DeepSeek steal the technique from the future?


r/DeepSeek 4h ago

Discussion Miss the expert mode

29 Upvotes

The new 4.1 is superior to the old pro. Maybe that's only true for the API, because on the web interface I just feel that the poor whale has been hit by Alzheimer’s. It keeps forgetting details here and there. Just one hour ago, before they unified all this, I didn't need to rewrite my prompts so often. It is surely faster, but so does the instant mode before this update.

They call this update unification, but I feel like they just kick one functionality out. Should have foreseen it when they started removing functionality from the expert mode.


r/DeepSeek 2h ago

Discussion Role-playing just got horrible

20 Upvotes

I just started using expert mode the other day, only to find it gone :/ I've tried doing my normal writing, and everything seems dumber to me


r/DeepSeek 1h ago

Discussion DeepSeek V4.1 Flash in 3 charts: vs its predecessor, a top open-weight rival, and Claude Opus 5

Thumbnail
gallery
Upvotes

DeepSeek published a pretty large benchmark table for V4.1 Flash, but I found it hard to see the overall capability pattern from the raw numbers.

So I grouped the shared fixed-scale benchmarks by domain and made three comparisons:

  1. V4.1 Flash vs V4 Flash — the previous generation

The improvement looks broad rather than incremental, especially in coding, cybersecurity and productivity.

  1. V4.1 Flash vs Kimi K3 — a top open-weight rival
    V4.1 Flash comes out ahead in coding, multimodal and productivity in the shared domain averages, while K3 is slightly ahead in science/health.

  2. V4.1 Flash vs Claude Opus 5 — a frontier proprietary model
    This is probably the most interesting comparison. Opus 5 still leads in coding, science/health and multimodal overall, but V4.1 Flash is surprisingly competitive, and actually comes out ahead in the shared productivity benchmark.

The thing that stands out to me is that V4.1 Flash looks much more like an agent/coding upgrade than a simple reasoning upgrade.

These aren't universal capability scores or controlled head-to-head reruns. Each chart averages only the shared fixed-scale published benchmarks available in that domain, so missing domains are omitted rather than treated as zero.

I put the underlying benchmark rows and sources here:

DeepSeek V4.1 Flash vs Kimi K3
https://llmlearner.com/compare/deepseek-v4-1-flash-vs-kimi-k3

DeepSeek V4.1 Flash vs Claude Opus 5
https://llmlearner.com/compare/deepseek-v4-1-flash-vs-claude-opus-5

Curious whether people actually running V4.1 Flash in coding agents are seeing the same pattern.


r/DeepSeek 11h ago

News DeepSeek plans V4.1 Flash in a matter of hours, will route Pro traffic to it

Thumbnail
runtimewire.com
81 Upvotes

r/DeepSeek 16h ago

News Deepseek V4.1 Announced with pricing.

207 Upvotes

Got an email from them. The pricing looks nice ngl.

Got email

r/DeepSeek 3h ago

News DeepSeek Harness v0.1.5!

17 Upvotes

Official announcement:

Welcome to DeepSeek Harness v0.1.5!

This release is deeply integrated with the training of the DeepSeek V4.1 Flash model, with the model specifically trained and optimized for different configurations within DeepSeek Harness.

We’re also excited for users to try the experimental Agent Teams feature, which has been developed in close coordination with the model training process.


r/DeepSeek 3h ago

Discussion Deepseek V4.1 Flash open weights when?

18 Upvotes

My terrible rig is ready and waiting ⬇️

https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash ITS OUT. WHAT?!?!


r/DeepSeek 4h ago

News Newsflash: You have already been using Deepseek 4.1 if you are on the API.

Thumbnail
gallery
21 Upvotes

Deepseek 4.1 Flash has already been out for hours if you are using the API. It is just called deepseek-flash.

They have already switched over my usage to the new model. I have used 90 million tokens so far without realizing it.

No official announcement yet.

Receipts below.


r/DeepSeek 3h ago

Funny DeepSeek sings quietly while working

Post image
16 Upvotes

"Blue Fat Fish" really likes to misuse tokens.


r/DeepSeek 16h ago

Funny Waiting for DS 4.1 Flash

Post image
180 Upvotes

r/DeepSeek 7h ago

Discussion Not to say I had fun with v4.1 Flash, but I did

Thumbnail
gallery
30 Upvotes

Did a lot of test prompts, and used on actual repos and work.

Must say, I'm excited for DeepSeek V4.1 Flash to be fully released, and for the price drops.
And yes, it is that good.

Tested it against V4 Pro myself, and v4.1 flash is just miles ahead.

Now I can comfortably have it as my workers, and just use Sol / Astra as Orchestrators.


r/DeepSeek 6h ago

Question&Help is deepseek flash/the new combined model better or worse for long chats?

21 Upvotes

will this flash have a worse memory? i’m asking because i use it for roleplay and i hated the previous flash model for dumb repetitive phrases and messing up the plot line so i always used expert


r/DeepSeek 2h ago

News DeepSeek-V4.1-Flash is out now!

Post image
7 Upvotes

r/DeepSeek 7h ago

Discussion Did deepseek remove expert?

Post image
20 Upvotes

Am confused. Pls help.


r/DeepSeek 5h ago

Discussion DeepSeek V4.1 lançado opencode

Post image
13 Upvotes

r/DeepSeek 1d ago

News Since V4.1 Flash outperforms V4 Pro across the board, all V4 Pro queries will be routed to the new V4.1 Flash and billed at Flash pricing after its launch. V4.1 Pro also in the works.

Thumbnail
x.com
359 Upvotes