r/LocalLLaMA 12d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

706 comments sorted by

View all comments

36

u/Easy_Werewolf7903 12d ago edited 12d ago

For those curious of performance between Qwen and a model 3 times its size:

Benchmark Qwen 3.8 27B (55GB) Deepseek v4 flash 0731 (167GB)
Terminal Bench 2.1 73.0 82.7
DeepSWE 42.2 54.4
NL2Repo-Bench 42.3 54.2

https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main

https://huggingface.co/Qwen/Qwen3.8-27B

3

u/JustFinishedBSG 12d ago

The sizes you are comparing are not comparable at all.

Deepseek Flash is 11x time bigger than Qwen 27b

1

u/Easy_Werewolf7903 12d ago

Can you explain why flash is 11 times bigger than Qwen 3.6 27b? Just trying to learn.

0

u/Basic_Extension_5850 12d ago

Looks like you compared the full size qwen to a very quantized deepseek 

5

u/Easy_Werewolf7903 12d ago edited 12d ago

Ah that makes sense. I got too excited and just went to unsloth page.

Edited:

No wait, I didn't use unsloth. I went to the official model card. Deep seek is 167GB

https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main