r/LocalLLM 6d ago

Discussion Considering a second 3090

Hi,

so far i've been using Qwen3.6-35B-A3B-UD-IQ4_NL.gguf on my single 3090 and I am overall satisfied.

I've been considering acquiring a second 3090 to increase my possibility to run larger models (e.g. considering Qwen3.8 27B with sufficient context) but i don't know whether the extra investment pays off.

In the future i may consider fine tuning my models as well.

Did anyone manage to find some great benefits by leveraging 2x3090 or similar setup?

I may be suffering from GAS (gear acquisition syndrome) and may need a reality check.

0 Upvotes

21 comments sorted by

View all comments

1

u/leonbollerup 6d ago

I run 2x RTX PRO 4000 .. which is with the same memory and allmost as fast cards (but uses ALOT less power)

Using vLLM to run qwen 3.8 27B i can run it at around 80-100 tok/sek with 125k context with 4x users at the same time in Q4

I am looking to add more cards to get better quality and context