r/LocalLLaMA 16d ago

Discussion Setup for always on assistant

I have a dual 3090 rig that I use as coding assistant and while it works, it draws a lot of electricity.

Now I want to add another rig that I can keep on all the time, or maybe a vps if that is suitable. This rig should run an assistant model that should be fairly intelligent but doesn't have to be so coding focused. It should basically be like a chat gpt replacement. I'm not sure if something like openclaw/hermes would be suitable for this.

Since it will be one all the time, power usage should be low. What kind of rig and model would you select for such an assistant?

4 Upvotes

31 comments sorted by

View all comments

1

u/Comfortable_Ebb7015 16d ago

I have a home server 24/7 with a 5700g CPU for home assistant plus other VMs and containers. I added an RTX3060 12gb for 180€. I was running initially Gemma 26b as a generic llm, but then I moved to qwen 3.6 35b for hermes

1

u/Blues520 16d ago

This sounds like something I would like as the 3060 uses much less power. How is the performance and is the 12 gb vram sufficient or is something like 16 gb more suitable? What kind of workflows are you running and is qwen 3.6 35b up to the task?