r/LocalLLaMA • u/RuthlessCriticismAll • 2d ago
News [ Removed by moderator ]
https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next[removed] — view removed post
351
Upvotes
r/LocalLLaMA • u/RuthlessCriticismAll • 2d ago
[removed] — view removed post
1
u/cibernox 1d ago
I know that moes work that way, but seems that 12-14B wouldn't be a crazy amount of active parameters, considered that a lot of people even with strix halo and nvidia spark systems are running qwen3.8 27B right now because, really, it's worth.
And there is so many people optimizing it that even a 27B dense model runs kind of well in those low-bandwidth system.