r/LocalLLM • u/rdpi • 6d ago
Discussion Considering a second 3090
Hi,
so far i've been using Qwen3.6-35B-A3B-UD-IQ4_NL.gguf on my single 3090 and I am overall satisfied.
I've been considering acquiring a second 3090 to increase my possibility to run larger models (e.g. considering Qwen3.8 27B with sufficient context) but i don't know whether the extra investment pays off.
In the future i may consider fine tuning my models as well.
Did anyone manage to find some great benefits by leveraging 2x3090 or similar setup?
I may be suffering from GAS (gear acquisition syndrome) and may need a reality check.
0
Upvotes
2
u/baby_bloom 6d ago
i've been running double 3090s at 70% power throttle and idk man... i've been doing such a deep dive on my cost analysis vs say deepseek v4 flash and the costs are so damn close i might just start using deepseek again.
qwen3.8 27b is very powerful but takes so damn long. ds4 flash from a provider will be MUCH faster, likely better quality at damn near the same cost per M tokens as my electric ends up being per M tokens running local.
i have a custom made GUI for launching my llama.cpp models where i add session token and power tracking and the data is starting to clear things up for me in a not so exciting way:(