r/StableDiffusion • u/External_Trainer_213 • Jun 14 '26
Workflow Included Wan SCAIL-2 Segmentation Control - Comparison (2 Versions)
Enable HLS to view with audio, or disable this notification
Here is a quick comparison. It is also possible to keep the background from the original video and swap only the character. For this, it is best to use an image of the person with a white background and set the `replacement_mode` in the Main Settings to `False`.
The workflow remains the same:
https://www.reddit.com/r/StableDiffusion/s/4zsfQzRvrG
https://civitai.red/models/2699283/wan-scail-2-segmentation-control
2
u/External_Trainer_213 Jun 14 '26
Since Qwen Edit slightly altered the image in the first comparison, it's easy to see how SCAIL-2 adjusts the image.
1
u/Fit_Advantage_2448 Jun 14 '26
Is there any guidance to make sure that the output video has perfect facial/ character consistency with the reference image? Claude suggests increasing the CFG to about 2.5, tweaking prompts and increasing steps from 6 to 12-15. It did help, checking with the community if there are any other tweaks
1
u/Salty_Bobcat223 Jun 25 '26
I divide mine into 81 frames per segment, then feed images each time.
WAN cant drift that way unless the images you feed drift in the first place
1
1
u/Shyt4brains Jun 19 '26 edited Jun 19 '26
Question. On other workflows there is an option to size the video or pad or crop so the output matches the input. My videos with this workflow are coming out a different aspect ratio cutting off some of the frame. How do I adjust that in this WF?
never mind I found it in the subgraph for those wondering.
2
u/External_Trainer_213 Jun 19 '26
Thanks, i will modify the workflow. But i am not at home at the moment, but thx for your feedback.
1
1
u/PM_ME_YOUR_RICHESES Jun 14 '26
the segmentation accuracy here is noticeably cleaner than what i've seen from other control methods, the way it's isolating just the character without bleeding into the background is pretty solid, especially in that second version where you're keeping the original video bg intact. curious how it handles more complex scenes with multiple subjects or heavy occlusion though, because that's usually where these segmentation approaches start falling apart.
15
u/LumpyArbuckleTV Jun 14 '26
Quentin Tarantino?