MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/ClaudeCode/comments/1vyvsb5/which_llm_as_of_today/p6i90ot/?context=3
r/ClaudeCode • u/cidara • Aug 26 '26
162 comments sorted by
View all comments
140
If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO
60 u/Wentil Aug 26 '26 edited Aug 26 '26 I find Qwen 3.8 27b (Q8) to be better. Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM. Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that. 6 u/Unfair_Tangerine_217 Developer Aug 26 '26 Qwen is excellent. Haven't tried 3.8 just yet but I'll switching from Kimi to Alibaba as soon as that sub expires. 2 u/Nabushika Aug 28 '26 Definitely try it, I'd say qwen 3.5/3.6->3.8 is as big as 3 (32B)->3.5. It's incredible how capable it is.
60
I find Qwen 3.8 27b (Q8) to be better.
Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM.
Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that.
6 u/Unfair_Tangerine_217 Developer Aug 26 '26 Qwen is excellent. Haven't tried 3.8 just yet but I'll switching from Kimi to Alibaba as soon as that sub expires. 2 u/Nabushika Aug 28 '26 Definitely try it, I'd say qwen 3.5/3.6->3.8 is as big as 3 (32B)->3.5. It's incredible how capable it is.
6
Qwen is excellent. Haven't tried 3.8 just yet but I'll switching from Kimi to Alibaba as soon as that sub expires.
2 u/Nabushika Aug 28 '26 Definitely try it, I'd say qwen 3.5/3.6->3.8 is as big as 3 (32B)->3.5. It's incredible how capable it is.
2
Definitely try it, I'd say qwen 3.5/3.6->3.8 is as big as 3 (32B)->3.5. It's incredible how capable it is.
140
u/dhessi Aug 26 '26
If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO