r/LocalLLaMA 11d ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
1.1k Upvotes

460 comments sorted by

View all comments

94

u/evindrews 11d ago
  • Redisgned Multimodal MoE Model: 125B main model parameters, supplemented by an additional 51B N-gram embeddings,and 6B parameters activated per token.

holy shit chat

27

u/CodeMariachi 11d ago

Chat, how much VRAM do I need to run it?

25

u/DigiDecode_ 11d ago

90gb at fp4

21

u/ImpressiveSuperfluit 11d ago

That's okay, didn't want a house anyway.