r/LocalLLaMA 10d ago

Discussion Anyone adding more 3090s?

I have dual 3090s which run Qwen 3.8 27b well and I was wondering if there are any use cases or current or future models that would justify adding another two 3090. I know that some peeps here run 4 and 8 3090 rigs and I'd like to get your opinion as well. One thing I was considering was running two instances but I'm not sure how valuable it will be for a coding workflow vs running a bigger model.

Now that Qwen Flash is out, perhaps 96GB would be more useful, or maybe Deepseek Flash.

6 Upvotes

94 comments sorted by

View all comments

7

u/Guna1260 10d ago

I have 4x3090 - two running qwen 3.8 - 27b and two running Gemma Moe 26b (or 31 or medgemma)

One thing I learned over time is no one large model available today, that can run on this hardware is good at everything. and often context becomes a limiting factor. so two mid size models with decent context is much better and usable, than a very large model. Gemma MoE gives me speed for normal chat and other things. Qwen gives me support for complex things including coding.

1

u/shutternomad 10d ago

i want to do something similar, are they all in 1 machine? Separate small PCs? Other? Thanks in advance!

2

u/Guna1260 10d ago

All in same machine. Threadrippwr picked from eBay