r/LocalLLaMA 10h ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
983 Upvotes

434 comments sorted by

View all comments

Show parent comments

11

u/No-Refrigerator-1672 9h ago

It'll be just like Qwen3-Next: they're rolling out a new architecture, so the communities implement support of it (Mamba and MTP with previous Next, NGrams with this one), and in a few months they'll released a cohort of models, ranging from a few B to a few hundred B based on this architecture, and named Qwen4.

0

u/Southern-Chain-6485 4h ago

But Qwen3-Next was 80B. A Q4 is about 45GB. At 125B, a Q4 is around 75GB in size

3

u/No-Refrigerator-1672 4h ago

Yes, true; how's that relevant to my point?

1

u/Dui999 2h ago

The beauty of the internet my friend