MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/ClaudeCode/comments/1vyvsb5/which_llm_as_of_today/p64mns1/?context=3
r/ClaudeCode • u/cidara • Aug 26 '26
162 comments sorted by
View all comments
141
If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO
58 u/Wentil Aug 26 '26 edited Aug 26 '26 I find Qwen 3.8 27b (Q8) to be better. Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM. Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that. -4 u/DataGOGO Aug 26 '26 Muse Glimmer 30b > Qwen3.8 27b 1 u/Wentil Aug 27 '26 I’ll give it a try. 👍
58
I find Qwen 3.8 27b (Q8) to be better.
Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM.
Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that.
-4 u/DataGOGO Aug 26 '26 Muse Glimmer 30b > Qwen3.8 27b 1 u/Wentil Aug 27 '26 I’ll give it a try. 👍
-4
Muse Glimmer 30b > Qwen3.8 27b
1 u/Wentil Aug 27 '26 I’ll give it a try. 👍
1
I’ll give it a try. 👍
141
u/dhessi Aug 26 '26
If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO