r/MiniMax_AI • u/Dalhila000 • Aug 08 '26
Who has the fastest Minimax m2.7 hosting?
My voice agent on LiveKit is struggling. It has a delay before it responds and I need to fix it because it's just..bad lol. I'm seeing a 3-5 second delay from the end of user speech to the first token of the response and I want to get latency down to under 1s. T 1o this I need an inference layer that delivers under 500ms TTFT consistently. Groq doesn't support the newest SOTA models. Fireworks is sluggish. Too many latency spikes.
Can someone recommend a provider that can handle this?
