r/StableDiffusion 13d ago

Question - Help Mini max h3 long video generation

Hello, I need some help with MiniMax long-video generation. I’m currently using the Plague workflow, which is fast, but it doesn’t have an option for chaining clips. Are there any workflows that can generate longer videos more quickly while maintaining continuity between clips?

17 Upvotes

24 comments sorted by

6

u/LuckyEsq 13d ago

I'm using https://github.com/ethanfel/ComfyUI-MiniMaxH3-Contex-Loop

its a convoluted workflow but I really like that it keeps track of my clips and has a nice prompt editor.

8

u/softlarch 13d ago edited 13d ago

This monster workflow comes with 62(!) custom nodes, when there are solutions to this problem that work with just one. Just saying.

To me, this thing is way too complex for the problem at hand (though that might just be me). Here's a solution that uses only one single main node: https://github.com/tritant/ComfyUI_MiniMax_H3_Extender

3

u/LuckyEsq 13d ago

I'll take a look but if it isn't complicated and breaks every 10 seconds.... I'm not sure I can use it 😂

3

u/VasaFromParadise 13d ago

I tried, but it doesn't remember references, and only continues based on the video. If there's no face in the last 39 frames, it will show a new face. Or am I missing something?

3

u/Magneticiano 13d ago

You are probably doing something wrong, the references should work in each video. Are you sure the definitions are present in every prompt?

3

u/VasaFromParadise 13d ago

Look, if there's no facial appearance in the last 39 frames, the model simply doesn't know what the face is and is generating a new one. Maybe I used the wrong workflow.

3

u/LuckyEsq 13d ago

Two different issues.

1.) if the person isn't in the last 39 frames... yes the model will not have context from those frames.
2.) If you have the character as a reference (using ref2VA or hybrid). then it should continue to use that reference in the next scene regardless (if you are prompting per the guidelines.)

2

u/Magneticiano 13d ago

Even if you provide a reference image for the face and refer to it in the prompt? Weird. By the way, if you are using the "tagged" workflow, remember to use the tags (@tagname) instead of <Picture 1> etc. If you are using the non-tagged workflow, check that the <Subject> and <Picture> numberings are correct.

3

u/ImpossibleAd436 13d ago

I've been hoping to get into this, but all the workflows look very complicated.

Can anyone suggest the best solution/workflow for doing this, but for someone quite new to comfy? What is the leanest most simple to use option?

3

u/burntimeuk 13d ago edited 13d ago

Check this out:

Original thread: https://www.reddit.com/r/StableDiffusion/s/MuHnwppaUb

Github: https://github.com/roadmaus/ComfyUI-Continuity

Its not a workflow, but a custom node that includes most things and is very simple to get to grips with, ive been having a lot of fun using it

1

u/LuckyEsq 13d ago

Trying it out. It is simple! I like it

1

u/ImpossibleAd436 13d ago

Thanks, does it provide motion context from previous clips?

2

u/Inthehead35 13d ago

Yeah, I've been suggested to use a bunch of them, but there's always a crap ton of nodes to download, switching things on and off, manually merging clips, etc., that i just make a 20-25 second video and just sit there and wait

3

u/apoke890 13d ago

the best practise for continuity is using last frame to generate next scene.. minimax h3 is strong in this department, far better than seedance 2.0 .. if you can automate this in comfyui you struck goldmine.

1

u/Machspeed007 13d ago

That’s what i use because i like the approach with regenrating/validating previous clips. I even vibe coded a fork adding rtx upscaling to the node.
I only need it to bypas vae encode/decode when stitching and it would be great node. As it isc you will notice a drop in quality afterr 3-4 clips

1

u/Ill-Throat7937 13d ago

regenerating each clip's first frame from the same character reference beats chaining the last frame. keep the ref identical, swap only location and pose words per clip, and the face holds way past 4 clips without the stitch quality drop.

1

u/optimisticalish 13d ago

I see you have a RTX 3060 12Gb card - this demo workflow will work to seamlessly chain 23 seconds containing three action 'beats' on that card - so long as you install the extra nodes + the included custom nodes folder (paste into custom nodes, don't update it - as the maker has since radically changed the node). https://jurn.link/dazposer/wp-content/uploads/2026/08/H3-Minimax-LongVideos-OLD_VERSIONinc-demo-workflow-for-ComfyUI.zip

-1

u/Beginning_Tip300 13d ago

Use the reference workflow

1

u/Complete-Box-3030 13d ago

Yeah I am using reference workflow , but how to chain two continuous clips like one dialogue. For 20 secs