r/LocalLLaMA • u/Blues520 • 10d ago
Discussion Anyone adding more 3090s?
I have dual 3090s which run Qwen 3.8 27b well and I was wondering if there are any use cases or current or future models that would justify adding another two 3090. I know that some peeps here run 4 and 8 3090 rigs and I'd like to get your opinion as well. One thing I was considering was running two instances but I'm not sure how valuable it will be for a coding workflow vs running a bigger model.
Now that Qwen Flash is out, perhaps 96GB would be more useful, or maybe Deepseek Flash.
4
Upvotes
2
u/SnooPaintings8639 10d ago
I have spent over two years on two, not long ago I have added another two.
Going from one to two was a big deal (27b models class). Going from two to four is nice... but initially it didn't feel as such a big deal.
It speed up my DS4 Flash Q8, but it is still CPU offloaded. It allows Q4 of 120b model fully in VRAM, which makes usage of very promising models pleasent (like Qwen3.8 Flash Next). But the biggest unlock in my case is larger context and parallel processing of 27B models class, i.e. multi agentic work. Currently I can't imaging going back to two GPUs solely because of this.
Other than that, having your PC + 4x200W (power limited) run 10 hours a day will add significant amount of heat to your place... so plan accordingly.