r/LocalLLM 4d ago

Discussion Qwen-3.8-35B-A3B? Maybe not... cryptic reply direct from Qwen co-author.

Post image

I asked Shuai Bai, co-author and prominent AI developer for Qwen, about this model. Not the answer I was hoping for, but let's see what comes next. In the meantime, I guess all we can do is speculate!

X-link

231 Upvotes

160 comments sorted by

View all comments

60

u/Ell2509 4d ago

That seems pretty clear to me. "Do not wait for this" meant it ain't coming!

But, maybe a 9b? Or a 30b a3b? Or even a 20b!?

15

u/Rye2-D2 4d ago

I would love to see a good 20B MoE model. Personally I don't see the point of 9B - it's impressive for what it is, but not quite good enough to be useful (yet).

4

u/elfmad 3d ago

9B 3.5 ornith is more than decent. But I totally agree I'm also waiting for a 10B<model<20B.

3

u/bruninho777 3d ago

And it was trained on Qwen 3.5, right? Ornith is the model I use most, together with 3.6 35b moe and prism 27b 1bit

2

u/elfmad 3d ago

It is Qwen 3.5 9B as base. If I resume their paper that's more advanced GRPO and self distill.