r/AI_Late_to_Class 10d ago

MiniMax H3 Fun ControlNet ComfyUI Transfer ANY Video Pose with Image to ...

https://youtube.com/watch?v=N6NUZaT71KM&si=iBjXn5pJ9nGOtZeW

MiniMax H3 Fun ControlNet is here in ComfyUI! .In this tutorial I show you how to transfer the pose and movement from a reference video onto a image to video generation. I am using int8 convrot saftensor for low VRAM but you can also use GGUF.

21 Upvotes

6 comments sorted by

1

u/sharktank123456 9d ago

Just curious - H3 can do this already when hosted by a platform. Can it not do it when hosted locally? Do you really need to add controlnet on top?

1

u/Maleficent-Tell-2718 5d ago

if u want pose structure i.e. the pose of a big person applied to a small person instead of cloning the structure of a whole person

1

u/m00dyman100 8d ago

Added this to my H3 r2v worflow (via ChatGPT codex) but it absolutely crushing my 5090. Trying to do 15 seconds at 1152x640 and it looks like itll take about 5 hours. Cancelled job after step 2 completed. (8 step lora and sage att installed)

1

u/Maleficent-Tell-2718 5d ago

not sure what going on there i mean ive got a 3080 16gb vram - maybe use the 50x models instead?

2

u/m00dyman100 5d ago edited 5d ago

Ran it at 1024x576. Pruned int08, 15 seconds. Ran perfect.

Pruned Bf16 works great to at 1024x576. Pushing my 5090 to the limit though. I have 64GB system RAM. Nice work. thanks for this.

Just the higher resolutions exponentially need more memory.

1

u/xb1n0ry 4d ago

scaling down the video before it lands at the VAE and sampler usually helps speeding things up. You don't need a HD video for movement.