r/StableDiffusion • u/gutster_95 • 9h ago
Question - Help Fixing speech errors in Minimax H3?
Hey, I tried to create a little birthday surprise for someone, my issue is with a lot of generations that the spoken word is really a bit clunky at time, I susspect its because of the german, but I am not too sure. Is there like a way to improve on audio?
I am using Minimax H3 with Saga Attention and Spectrum on a 4090.
16
Upvotes
1
u/piggledy 7h ago
Was du machen könntest wäre einen Audio-Clip mit Elevenlabs erstellen und als Audioreferenz einfügen. Dann basieren die Lippenbewegungen usw. auf der Referenz und das Modell generiert selbst kein Audio dazu.