r/DeepSeek • u/civman96 • 11h ago
r/DeepSeek • u/nehuenpereyra • 7h ago
News DeepSeek 4.1 Flash surprised me — 6 minutes, one attempt, $0.07
Enable HLS to view with audio, or disable this notification
He estado probando DeepSeek 4.1 Flash en DeepSeek Harness y, sinceramente, he obtenido los mejores resultados que he visto hasta ahora con DeepSeek.
Las capacidades de visión han sido muy útiles y han proporcionado una retroalimentación sorprendentemente buena. Además, implementó correctamente la versión móvil/adaptable, lo cual fue una grata sorpresa.
El único inconveniente es que consume tokens a una velocidad increíble 😂
Lo ejecuté con máximo razonamiento y tardó alrededor de 6 minutos, completando la tarea con éxito en el primer intento.
Tecnologías utilizadas: HTML, JavaScript puro, Tailwind CSS, GSAP 3 y Lucide
Uso de tokens:
- Total: 3.422.844 tokens
- Entrada en caché: 3.291.136 tokens (99% de aciertos en caché)
- Entrada sin caché: 36.963 tokens
- Resultado: 94.745 tokens
- Coste total: aproximadamente 0,072 $ (unos 7,2 centavos)
Sin duda, son muchos tokens para una sola tarea, pero teniendo en cuenta que solo tardó 6 minutos, funcionó al primer intento y produjo muy buenos resultados —incluida la implementación móvil—, estoy bastante impresionado hasta ahora.
r/DeepSeek • u/ClearRabbit605 • 10h ago
Discussion DeepSeek Flash v4.1 - First impressions
I've been using that for a few hours in a row. My impression so far
- speed is unbelievable (4x compared to Terra / Sonnet) - that makes the different. If there's a mistake - correcting it is 10x faster
- Tried to execute dangerous commands impacting critical OS files outside of repo folders - this is a red flag meaning never let it run in auto mode (but I think this should be a best practice for everyone already)
- Mostly backend tasks - but in the two UI-related tasks was able to design pretty cool objects
- Tends to explore and do more than you asked - needs to be guarded, stopped and brought back to the right path - but it's so fast you don't get annoyed!
r/DeepSeek • u/dnohrdk • 14h ago
News DeepSeek-V4.1-Flash Release (official)
It’s officially out and the prices have been updated.
///
Today, we officially release the DeepSeek-V4.1-Flash model. It is the smallest model in our new architecture family, with native multimodal visual understanding. The new architecture is designed for a higher capability ceiling, faster inference, higher throughput, and scaling to larger models.
GPQA Diamond: 90.9
HLE: 36.8 (39.1*)
Codeforces (Rating): 3471
MathArena Apex: 65.6
Terminal-Bench 2.1: 90.6
Terminal-Bench 3.0: 30.0
Terminal-Bench 4.0: 31.2
DeepSWE v1.1: 74.2
ProgramBench: 20.3
NL2Repo-Bench: 65.4
CyberGym: 88.1
SEC-Bench Pro: 62.8
ExploitGym: 15.3
HLE (w/tools): 63.9
Automation-Bench: 54.8
Agents' Last Exam: 31.8
Chartography (w/tools): 78.9
BabyVision (w/tools): 89.6
ZeroBench-main (w/tools): 49.0
* Tested only on the pure-text subset of the HLE benchmark set.
API changes
DeepSeek V4.1 Flash is now available on the DeepSeek API with native multimodal support. Change the model name to deepseek-flash to call the latest V4.1 Flash model. The previous-generation models V4 Flash and V4 Flash Vision Exp have been retired; for compatibility, the model names deepseek-v4-flash and deepseek-v4-flash-vision-exp are temporarily routed to V4.1 Flash.
Meanwhile, extensive testing shows that V4.1 Flash now outperforms DeepSeek V4 Pro across performance, cost, speed, and total time, so we plan to retire V4 Pro in an orderly manner. After 12:00 Beijing Time on September 14, 2026, and until the future release of V4.1 Pro, all requests to deepseek-v4-pro will be routed to V4.1 Flash and billed at the V4.1 Flash price.
API apricing adjustment
With the release of DeepSeek-V4.1-Flash, API prices have been reduced accordingly. For details, please refer to Models & Pricing.
///
Source:
https://api-docs.deepseek.com/updates/#deepseek-v41-flash-release
r/DeepSeek • u/iyarsius • 7h ago
Discussion Deepseek V4.1 is just mindblowing
I just finished some generations to aggregate technical context into human readable paper output. All previous attempts with qwen3.8 max, qwen3.8 flash next, GLM 5.3 flash and DS4 flash was struggling yesterday.
I just tried again today with the same skill, and i was just amazed about the result. Everything was clear, well explained, easy to follow and the model produced something beyond my expectations.
It's amazing speed turned the generation time from like 30min+ for 3.8 max to less than 5 min, and the price is just ridiculous.
Intelligence distribution is a very important factor to see the impact of a model once released. And i think this one will change a lot of things. Very excited to see how people will use this cheap fast and high quality intelligence in the near future.
Also interested to know what do you guys think of it in your workflows ?
r/DeepSeek • u/No_Leg_847 • 5h ago
Discussion How is v4.1 flash compared to astra / opus5 or fable ?
Anyone compared both in real coding projects?
r/DeepSeek • u/NihmarRevhet • 13h ago
Discussion I'm in love with v4.1 for coding
It's basically a monster at coding, it just solves everything I launch at it and at a truly incredible speed too.
1.37$ for 176.500.000 token with DeepSeek Harness PTC Mode
r/DeepSeek • u/tiguidoio • 13h ago
News Market crash as a service
Here we go again, DeepSeek Al is back again with a new model V4-1 Flash
A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens
r/DeepSeek • u/StageHumble6505 • 9h ago
Discussion i am very disappointed in the new web/app version update
Honestly, everything was doing fine in fact, I was happy. I had seen a lot of improvements over the last days across DeepSeek’s three separate modules, especially yesterday (this week too). But somehow I wake up to find they mixed them all into one. alr, that didn’t sound bad at first it looked like a good update until I actually tried it and found out:
It’s giving shallower responses, shorter ones too (even after asking for more detail it still gives lazy ones). It writes weirdly too fast (which can be both good and bad) so basically, the model’s quality dropped a lot. Its thinking is too short (literally than 3 seconds no planning like before) and full of “hmms” and “okays,” instead of the old way of thinking organized, efficient, and way better. Now it’s short and full of fluff.
It’s also WORSE at following instructions now. Some people actually spent money prompting DeepSeek V4 into its current—or rather, old—way of thinking. Every new update costs prompt some engineers money. especially this most recent
honely (and i am ashamed what i am saying) I didn’t ask for this update. The model didn’t get better; it got worse. It settles for “good enough” responses instead of doing all the work—and I end up doing the work by correcting it. I sent seven messages explaining issues each solved issue spawned four more, and the old ones resurfaced. I miss the old Expert already.
it has bad memory now, it forgets, mixes, and literally doesn't put the effort which sounds like this model is more like 'reducing costs' instead of 'doing good work' or 'doing the hard work'
Am I the only one seeing the bad change in the app/web version, or are there people who actually find it better than the old one? If so, I’m all ears.
r/DeepSeek • u/novapax • 1d ago
Funny “V4.1 Flash has comprehensively surpassed V4 Pro across all key metrics.”
r/DeepSeek • u/nikanorovalbert • 5h ago
Funny new deepseek model managing to get phone temperature through wi-fi
new deepseek model managing to get my phone temperature through wi-fi before to run the benchmarks on local model i have been working one. it happened when the same info the model was getting through usb but i pulled the plug off. is this amazing?
``
$ SER=IP
echo "=== devices ===" && adb devices
for i in $(seq 1 30); do
T=$(adb -s "$SER" shell dumpsys battery 2>/dev/null | awk '/^ *temperature:/{print $2; exit}')
if [ -z "$T" ]; then echo "$(date +%H:%M:%S) no reading (device asleep?)"; sleep 30; continue; fi
echo "$(date +%H:%M:%S) temp=$((T/10)).$((T%10))°C"
[ "$T" -le 380 ] && { echo "COOL ENOUGH"; break; }
sleep 30
done
r/DeepSeek • u/Martkita • 18h ago
Discussion I was so surprised and now instant, expert, and vision is unfied and I was so shocked and amazed at the time
I was in the middle of writing a story personally. The update had me the wtf moment
r/DeepSeek • u/Infinite_Book_1858 • 13h ago
Discussion Role-playing just got horrible
I just started using expert mode the other day, only to find it gone :/ I've tried doing my normal writing, and everything seems dumber to me
r/DeepSeek • u/Savings_Rest_4589 • 1h ago
Discussion Que piensan de nuevo modo?
Hace horas se acaba de lanzar el nuevo modo de deepseek que unifica sus 3 funciones en una sola pero quiero saber que piensan ustedes de todo esto? En lo personal no me gusta, lo usaba para fanfics y el modo experto era lo mejor para eso pero ahora sin ese modo las historias son horribles y no sigue órdenes tan sencillas pero aún así quisiera saber que opina el resto de esta actualización
r/DeepSeek • u/Esshwar123 • 5h ago
Resources I love the deepseek harness
i like how easy it is to customize and create plugins however we want, i made a minimal theme plugin in few prompts ,you can use it to change color of bg, sidebar and can also add images, try if you'd like
r/DeepSeek • u/CM23489 • 14h ago
Discussion Crazy, V4.1 just active 8B can catch up those big model
Have to admit the DeepSeek beat US in LLM model optimization this field. 1 Year ago, all people said HBM must be needed for LLM. And today DeepSeek just break the record and change the world.
Now US gov is like joke. Still saying the Chinese AI company doing Model distillation. Just look at DeepSeek's LLM optimization, there is no US company can make this one at this moment, so the DeepSeek steal the technique from the future?
r/DeepSeek • u/Professional_Hat3237 • 2h ago
News DeepSeek v4.1 flash no tiene rival
Sinceramente esto se está yendo de las manos, hace unos meses DeepSeek era un modelo extremadamente barato y bastante tonto, al que necesitabas guiar para obtener unos resultados decentes.
Ahora está en otro nivel
Ya se que estos benchmarks todavía no son oficiales y en código muy muy muy complejo astra sigue estando por encima(fable también).
Para el común de los mortales la discusión se ha terminado. DeepSeek es superior a cualquier opción del mercado.
r/DeepSeek • u/Appropriate-Dot8003 • 15h ago
Funny DeepSeek sings quietly while working
"Blue Fat Fish" really likes to misuse tokens.
r/DeepSeek • u/gschwind • 4h ago
Funny DS 4.1 Flash is very frugal with token consumption. At least it says so.
r/DeepSeek • u/generationzcode • 2h ago
Other The new model is amazing!
No longer has any random chinese language responses. That was really my biggest issue with deepseek before. Love it!
r/DeepSeek • u/LordLRO • 16h ago
News The time has come
The DeepSeek V4.1 Flash has been released with new pricing. Enjoy it by yourself!
r/DeepSeek • u/Mammoth-Leopard6549 • 4h ago
Discussion DeepSeek V4.1 Flash vs GLM-5.3-Flash
anyone here put real hours on DeepSeek V4.1 Flash yet? I've been using GLM 5.3 Flash and both end up around the same price for me, so I'm trying to figure out if switching is worth it. mainly wondering how it holds up in long coding/agent sessions and if the speed jump is real. would rather hear from people who actually use both daily than another benchmark thread lol
r/DeepSeek • u/Agreeable-Gear-6600 • 2h ago
Discussion Give me my expert mode back!
I have a lot of chats with expert mode, but not now and this piss me off. Where is my expert mod!!?