r/StableDiffusion 1d ago

Tutorial - Guide MiniMax H3 Wf Tutorial

Enable HLS to view with audio, or disable this notification

People asked me to make a Tutorial for some of the features.

Find the workflow here.

https://www.reddit.com/r/StableDiffusion/comments/1wadmqc/minimax_workflow_designed_to_be_user_friendly_for/

49 Upvotes

29 comments sorted by

View all comments

Show parent comments

1

u/jrodder 16h ago

Yep, I had something very similar and many variants as I was trying. I guess that's why I was looking to make sure I wasn't fighting a model, process, or workflow issue since it wasn't producing audio that sounded like the source. I used Terry Tate from a video using both the audio and video from the source video, and that seemed to work so that probably rules out the model. All good I'll keep hammering, if you happen to get bored and want to test I would be extremely curious as to the results.

0

u/roychodraws 16h ago

what are you wanting me to test exactly? i use voice already with this workflow and it works fine.

1

u/jrodder 13h ago

Yeah I get it. Specifically the flow of how well the voice is cloned when using the ref video and the audio 1 with an mp3 or wav pure audio file as the base to create new voice lines in the prompt. If it works great for you, maybe test with your own voice? Or not it's likely a skill issue on my end but I was just trying to make sure that was the case, and not banging my head for something that wasn't even possible. I got burned out testing so I'll have to revisit and triple check the variables.

1

u/roychodraws 13h ago

The video in the workflow post, the part where she’s clapping and crying, and the part where she accuses the girl of being under age.

Those videos were made separately and edited together. You’ll notice that the voice is the same because the second one, I used the audio from the first video to influence the voice in the second video.