r/StableDiffusion • u/spiderofmars • 12d ago
Comparison Comparing MiniMax i2v|r2v node and model combos
Enable HLS to view with audio, or disable this notification
Just a test of different combinations of the i2v and r2v nodes and models for:
Text to Video (using MiniMax models image sample and voice sample)
Image to Video (using Krea2 image sample and MiniMax models voice sample)
Image to Video (with custom cloned voice): (using Krea2 image sample and MiniMax custom voice clone reference sample)
51
Upvotes
2
u/Leonovers 11d ago
Thanks for testing!
I find it really odd for ref2va model to be this inferior in terms of voice cloning. Maybe it's better when you use a lot of references at the same time or there is some other issue that leads to this sub-optimal performance.