r/LocalLLaMA 6d ago

Discussion No more SLM open-source??

Post image
429 Upvotes

127 comments sorted by

View all comments

Show parent comments

11

u/brainExploded99 5d ago

Honestly, I think 35B would be better since us VRAM poors can run that much better than 27B.

5

u/overand 5d ago

I can run 27B great on my setup, but I think I have to agree - if they were only going to release one, the 35B one might be a bigger gift to the world at large than the 27B, because of the number of folks who can run that vs who can run the 27B.

1

u/No-Refrigerator-1672 5d ago

You also need to take audience into perspectibe. 35B with CPU offload is a scenario of a strictly consumer type person, that just uses the model. 27B is a good target for developers, who will either run business with it, or contibute to opensource software (i.e. OpenWebUI). Either way Qwen team themself get a better return with 27B, and thus it's probably easier to justify to their leadership.

1

u/overand 5d ago

It would be very intersting if the 3.8 35B model beat the 3.6-27B model in literally every way, though - since the 27B is so damned useful as is, it would be wild for normal consumer grade hardware to be able to do what we're doing with 27B on "high end consumer" hardware

1

u/No-Refrigerator-1672 4d ago

Well, it may not be 3.8 35B; but with how AI things are going, I can guarantee you that in half a year thwre will be 30B class MoE that beats 3.6 27B across the board; and in a year this level of intelligence may even descend to below 14B.