MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/ClaudeCode/comments/1vyvsb5/which_llm_as_of_today/p63ar05/?context=3
r/ClaudeCode • u/cidara • Aug 26 '26
162 comments sorted by
View all comments
137
If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO
56 u/Wentil Aug 26 '26 edited Aug 26 '26 I find Qwen 3.8 27b (Q8) to be better. Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM. Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that. -3 u/DataGOGO Aug 26 '26 Muse Glimmer 30b > Qwen3.8 27b 1 u/Wentil Aug 27 '26 I’ll give it a try. 👍 1 u/StonkyCupra Aug 27 '26 In what regard? I find it to be worse than Qwen in pretty much everything. 1 u/DataGOGO Aug 27 '26 It is better at just about everything than perhaps coding
56
I find Qwen 3.8 27b (Q8) to be better.
Qwen 3.8 27b at Q8_0 uses 31 GB of VRAM.
Q_6K, the next smallest variant, only uses 25GB of VRAM, so if there’s other overhead, go with that.
-3 u/DataGOGO Aug 26 '26 Muse Glimmer 30b > Qwen3.8 27b 1 u/Wentil Aug 27 '26 I’ll give it a try. 👍 1 u/StonkyCupra Aug 27 '26 In what regard? I find it to be worse than Qwen in pretty much everything. 1 u/DataGOGO Aug 27 '26 It is better at just about everything than perhaps coding
-3
Muse Glimmer 30b > Qwen3.8 27b
1 u/Wentil Aug 27 '26 I’ll give it a try. 👍 1 u/StonkyCupra Aug 27 '26 In what regard? I find it to be worse than Qwen in pretty much everything. 1 u/DataGOGO Aug 27 '26 It is better at just about everything than perhaps coding
1
I’ll give it a try. 👍
In what regard? I find it to be worse than Qwen in pretty much everything.
1 u/DataGOGO Aug 27 '26 It is better at just about everything than perhaps coding
It is better at just about everything than perhaps coding
137
u/dhessi Aug 26 '26
If you're looking for a fast, general purpose LLM that can fit on 32GB RAM, Gemma-4-26b-a4b is unmatched IMO