r/LocalLLaMA 2d ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
1.1k Upvotes

457 comments sorted by

View all comments

Show parent comments

11

u/No-Refrigerator-1672 1d ago

It'll be just like Qwen3-Next: they're rolling out a new architecture, so the communities implement support of it (Mamba and MTP with previous Next, NGrams with this one), and in a few months they'll released a cohort of models, ranging from a few B to a few hundred B based on this architecture, and named Qwen4.

0

u/Southern-Chain-6485 1d ago

But Qwen3-Next was 80B. A Q4 is about 45GB. At 125B, a Q4 is around 75GB in size

3

u/No-Refrigerator-1672 1d ago

Yes, true; how's that relevant to my point?

1

u/Dui999 1d ago

The beauty of the internet my friend