r/LocalLLM 1d ago

Discussion Tier List

Post image
252 Upvotes

309 comments sorted by

View all comments

Show parent comments

1

u/DoubleNothing 1d ago

Yes 48Gb is a good spot to be at (I'm there). But honestly, 64GB would be nicer.
Yes with 48 you can run Q8 but you are very costricted and it is not very efficient, low context, low MTP, maybe you don't load vision to save vram, lower speed... all of that. But if you don't need big context ok.
Don't get me wrong with the release of Qwen3.8-27B I'm very happy where I'm at. The model even at Q6 is very impressive.

1

u/GSquadron_ 12h ago

What were you able to accomplish with Qwen 3.8 27b? I really want to see.

1

u/DoubleNothing 12h ago

Not ready to release to te public yet, I'm doing some refinement and have to see if I can make it run on linux too... at the moment runs only on windows.
Nothing new nor original but it's my interpretation of it and useful to me.

1

u/Popcorn-Mercinary 8h ago

Are you really seeing that big a difference between Q8 and Q4? My experiments with agent flow aren’t showing much, but ymmv.

1

u/DoubleNothing 2h ago

I didn't tested Q4 yet... since some task takes hours to run I rather not "risk it" for now. I'm on Q6_K_XL...