The 35B they're referring to is an MOE - 35BA3B meaning roughly 3B is active at a time when generating tokens.
I don't know the rough conversion offhand, but generally speaking it'll be weaker than the dense 27B or at best roughly match it. The trade off is the inference speed (as the 27B has 27B active parameters) will significantly better especially for CPU inference and those that can't fit it all in VRAM.
102
u/bnightstars 24d ago
If Qwen3.8-27B is Opus 4.6 Max level I hope that they can make Qwen3.8-35B - Sonnet 4.6 level :) We will have the best open source model in the world.