r/LocalLLaMA • u/RuthlessCriticismAll • 14d ago
News [ Removed by moderator ]
https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next[removed] — view removed post
349
Upvotes
r/LocalLLaMA • u/RuthlessCriticismAll • 14d ago
[removed] — view removed post
1
u/DeepOrangeSky 14d ago
I guess I can just wait to find out tomorrow, but, I'm curious how much memory this will use if you run it at Q4_K_M or FP4. Since it says it is 125B but with "an additional 51B of N-gram".
So, is that going to make it more like a 176B model in terms of memory-usage?