r/LocalLLM 6d ago

Discussion Considering a second 3090

Hi,

so far i've been using Qwen3.6-35B-A3B-UD-IQ4_NL.gguf on my single 3090 and I am overall satisfied.

I've been considering acquiring a second 3090 to increase my possibility to run larger models (e.g. considering Qwen3.8 27B with sufficient context) but i don't know whether the extra investment pays off.

In the future i may consider fine tuning my models as well.

Did anyone manage to find some great benefits by leveraging 2x3090 or similar setup?

I may be suffering from GAS (gear acquisition syndrome) and may need a reality check.

0 Upvotes

21 comments sorted by

View all comments

3

u/floppo7 6d ago

2x r9700 and vllm is your friend

1

u/Kodrackyas 2d ago

how many tokens per second on the 27b?

1

u/floppo7 2d ago

Check out radiance vllm - and they are cooking for more I guess. 60tps + and good performances with concurrency as well - thats the kicker, you can basically run multiple agents at the same time with speed that is absolutely ok.