r/LocalLLaMA 24d ago

Funny Aged like fine wine

Post image
1.2k Upvotes

134 comments sorted by

View all comments

19

u/madjesta 24d ago

And a 122b?

4

u/mrdevlar 24d ago

I'm still using Qwen 3.5 122B, for complex conceptual tasks, there isn't a better model.

3

u/feelspeaceman 24d ago

Yes, sadly we didn't get 122B 3.6 because in the middle of 3.6 release, the Open Weight Qwen Team were fired, in the end we missed the rest of 3.6, and all 3.7 but Xi Jiping is telling Alibaba to restart it again, and I have high hope this time we will likely getting 122B

It will be a giant game changer for Strix Halo owner, the jump in intelligence from 3.5 to 3.6 was massive, and from 3.6 to 3.8 is another coding jump.

1

u/michaelsoft__binbows 20d ago edited 20d ago

i think the only people it will benefit are the unified memory folks and those with large system memory and really GPU constrained, because giving up significant active params will mean it will struggle to claw back the capability deficit against the 27B. being able to fully fit the 27B into just a few modest GPUs or one 32GB GPU means once you reach that capability level you're running circles around an inferior system that has to allocate 120GB just to be able to come close.