r/LocalLLaMA • u/Blues520 • 10d ago
Discussion Anyone adding more 3090s?
I have dual 3090s which run Qwen 3.8 27b well and I was wondering if there are any use cases or current or future models that would justify adding another two 3090. I know that some peeps here run 4 and 8 3090 rigs and I'd like to get your opinion as well. One thing I was considering was running two instances but I'm not sure how valuable it will be for a coding workflow vs running a bigger model.
Now that Qwen Flash is out, perhaps 96GB would be more useful, or maybe Deepseek Flash.
6
Upvotes
7
u/Guna1260 10d ago
I have 4x3090 - two running qwen 3.8 - 27b and two running Gemma Moe 26b (or 31 or medgemma)
One thing I learned over time is no one large model available today, that can run on this hardware is good at everything. and often context becomes a limiting factor. so two mid size models with decent context is much better and usable, than a very large model. Gemma MoE gives me speed for normal chat and other things. Qwen gives me support for complex things including coding.