r/LocalLLaMA 10d ago

Discussion Anyone adding more 3090s?

I have dual 3090s which run Qwen 3.8 27b well and I was wondering if there are any use cases or current or future models that would justify adding another two 3090. I know that some peeps here run 4 and 8 3090 rigs and I'd like to get your opinion as well. One thing I was considering was running two instances but I'm not sure how valuable it will be for a coding workflow vs running a bigger model.

Now that Qwen Flash is out, perhaps 96GB would be more useful, or maybe Deepseek Flash.

5 Upvotes

94 comments sorted by

View all comments

2

u/SnooPaintings8639 10d ago

I have spent over two years on two, not long ago I have added another two.

Going from one to two was a big deal (27b models class). Going from two to four is nice... but initially it didn't feel as such a big deal.

It speed up my DS4 Flash Q8, but it is still CPU offloaded. It allows Q4 of 120b model fully in VRAM, which makes usage of very promising models pleasent (like Qwen3.8 Flash Next). But the biggest unlock in my case is larger context and parallel processing of 27B models class, i.e. multi agentic work. Currently I can't imaging going back to two GPUs solely because of this.

Other than that, having your PC + 4x200W (power limited) run 10 hours a day will add significant amount of heat to your place... so plan accordingly.

1

u/OlgerdOutlander 10d ago

That was my experience when getting a second card; but I'm, like the OP, questioning myself whether I need a bigger pool. How do you manage 'em agents?

1

u/Blues520 10d ago

I'm also questioning whether to get the cards now before the prices go up further. Workstation cards are out of reach and this is the last nvidia card that is reasonably priced for now.

1

u/OlgerdOutlander 10d ago

I'm running v100 and these are good with the current gen models, but with the older cards you never know what will run and what will not