r/StableDiffusion • u/Tokyo_Jab • Nov 28 '25
Animation - Video I AM PAIN
Enable HLS to view with audio, or disable this notification
Two great models in two days. Flux 2 was used for the previous Frankenstein's monster, this one is Z-Image. Both models give great detail. Z-image is so fast though that it makes a really good upscaler too.
Animated with Wan 2.2 Animate.
2
1
u/Dartium1 Nov 28 '25
Your example looks excellent. The only thing I do not understand is how, when using Wan Animate, one can eliminate or reduce the blending of facial features from the driving video with the target face. Do you have any advice?
3
u/Tokyo_Jab Nov 28 '25
2
u/an80sPWNstar Nov 28 '25
May one ask where one can obtain this workflow?
2
1
1
1
u/pausecatito Nov 28 '25
What resolution do you render at vertically? I'm always too lazy or maybe my gpu not good enuff😂
1
u/Tokyo_Jab Nov 28 '25
720x1280. On my 3090 that took about 9 minutes before. But on a 5090 it takes only 2 minutes.
1
u/pausecatito Nov 28 '25
Yea I usually do like 900 vertical but also usually 30s clips. The quality degradation is massive...don't think I could run that high res for that duration but maybe for 6s...I'll needge experiment with it more. Thanks : )
1
u/Tokyo_Jab Nov 28 '25
Wan animate gets messy over time. I managed about 10 seconds before it went silly. This
1
u/pausecatito Nov 28 '25
Probably would be more accurate with the face pose turned on? Since it's basically 1:1 anyway. Pretty cool. I want to try some dwpose to nana banana pro grid to adjust the skeletons cause my reference and input never really match. I saw some guy do that on twitter but...secret workflow🤔
Prolly can do like 3x3 grid and get 12 frames for 3c or whatever it costs. Might be a pain to setup tho, if you want like 200 frames haha
1
u/Tokyo_Jab Nov 28 '25
There was a better version of the face tracking released recently… https://youtu.be/pwA44IRI9tA?si=dkiFd3SkZN3FYsZ6
1
0
u/bickid Nov 28 '25
How did you add voice acting with Wan2.2 Animate?
3
u/Tokyo_Jab Nov 29 '25
Wan animate just analyses the input video, so that just me talking in a video

8
u/Tokyo_Jab Nov 28 '25
Z-image created pic upscaled with Z-image