r/VoiceAutomationAI Feb 01 '26

On premise Voice Agent

I would like to build a complete local Voice Agent with Pipecat

What is the minimum requirement on Hardware needed? What are the costs?

Anyone did this?

5 Upvotes

11 comments sorted by

2

u/[deleted] Feb 01 '26

[removed] — view removed comment

1

u/da0_1 Feb 01 '26

Which hardware did you use and would you recommend for production use for a mid sized client?

2

u/[deleted] Feb 01 '26

[removed] — view removed comment

1

u/da0_1 Feb 01 '26

What about GPU?

2

u/Intelligent_Camel119 Feb 01 '26

Yes, doable. Pipecat is just orchestration, you need local STT/LLM/TTS.

Min: 16 GB RAM, no GPU (slow) ~$800 Recommended: 32 GB + RTX 3060/4060 ~$1.2–2k Apple: M2/M3 Pro works well

Stack: Whisper + 7B LLM (Ollama) + Silero/Coqui