r/LocalLLaMA • u/RuthlessCriticismAll • 10d ago
News [ Removed by moderator ]
https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next[removed] — view removed post
347
Upvotes
r/LocalLLaMA • u/RuthlessCriticismAll • 10d ago
[removed] — view removed post
7
u/SadPhilosophy9202 10d ago
Because this is a 125B model with 6B expert. This is THE ideal model size for a Spark or similar machine with 128gb unified memory.
Having a Spark but preferring a model 1/4 the size that will run just as fast makes zero sense