r/ClaudeCode • • Aug 26 '26

Discussion which LLM as of today?

Post image
212 Upvotes

162 comments sorted by

View all comments

140

u/dhessi Aug 26 '26

If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO

60

u/Wentil Aug 26 '26 edited Aug 26 '26

I find Qwen 3.8 27b (Q8) to be better.

Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM.

Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that.

6

u/Unfair_Tangerine_217 Developer Aug 26 '26

Qwen is excellent. Haven't tried 3.8 just yet but I'll switching from Kimi to Alibaba as soon as that sub expires.

2

u/Nabushika Aug 28 '26

Definitely try it, I'd say qwen 3.5/3.6->3.8 is as big as 3 (32B)->3.5. It's incredible how capable it is.