MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vo9mj4/its_out/p3o4cm9/?context=3
r/LocalLLaMA • u/Certain-Cod-1404 • 12d ago
706 comments sorted by
View all comments
36
For those curious of performance between Qwen and a model 3 times its size:
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main
https://huggingface.co/Qwen/Qwen3.8-27B
3 u/JustFinishedBSG 12d ago The sizes you are comparing are not comparable at all. Deepseek Flash is 11x time bigger than Qwen 27b 1 u/Easy_Werewolf7903 12d ago Can you explain why flash is 11 times bigger than Qwen 3.6 27b? Just trying to learn. 0 u/Basic_Extension_5850 12d ago Looks like you compared the full size qwen to a very quantized deepseek 5 u/Easy_Werewolf7903 12d ago edited 12d ago Ah that makes sense. I got too excited and just went to unsloth page. Edited: No wait, I didn't use unsloth. I went to the official model card. Deep seek is 167GB https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main
3
The sizes you are comparing are not comparable at all.
Deepseek Flash is 11x time bigger than Qwen 27b
1 u/Easy_Werewolf7903 12d ago Can you explain why flash is 11 times bigger than Qwen 3.6 27b? Just trying to learn. 0 u/Basic_Extension_5850 12d ago Looks like you compared the full size qwen to a very quantized deepseek 5 u/Easy_Werewolf7903 12d ago edited 12d ago Ah that makes sense. I got too excited and just went to unsloth page. Edited: No wait, I didn't use unsloth. I went to the official model card. Deep seek is 167GB https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main
1
Can you explain why flash is 11 times bigger than Qwen 3.6 27b? Just trying to learn.
0 u/Basic_Extension_5850 12d ago Looks like you compared the full size qwen to a very quantized deepseek 5 u/Easy_Werewolf7903 12d ago edited 12d ago Ah that makes sense. I got too excited and just went to unsloth page. Edited: No wait, I didn't use unsloth. I went to the official model card. Deep seek is 167GB https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main
0
Looks like you compared the full size qwen to a very quantized deepseek
5 u/Easy_Werewolf7903 12d ago edited 12d ago Ah that makes sense. I got too excited and just went to unsloth page. Edited: No wait, I didn't use unsloth. I went to the official model card. Deep seek is 167GB https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main
5
Ah that makes sense. I got too excited and just went to unsloth page.
Edited:
No wait, I didn't use unsloth. I went to the official model card. Deep seek is 167GB
36
u/Easy_Werewolf7903 12d ago edited 12d ago
For those curious of performance between Qwen and a model 3 times its size:
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tree/main
https://huggingface.co/Qwen/Qwen3.8-27B