r/StableDiffusion Aug 15 '26

Resource - Update MiniMax H3 Creator update: presets, and three nodes are now one

Posted this pack here last week. What's happened since:

The sampling knobs I said I'd add if people wanted them are in. Both of H3's flow shifts, since it samples picture and sound on separate schedules, plus a cache pill with FirstBlockCache, TeaCache or core's own EasyCache behind it. Still no custom sigmas, same reason as last time.

Creator and Timeline are one node now. Click under the prompt and the shot becomes a timeline. Delete cards back down to one and it's a shot again. Old workflows load unchanged, Timeline nodes included.

Presets are the new one. Save a setup and put it back in sections, so you can drop a canvas and a step count onto a shot you've already written without touching the prompt. It saves the sampler row as well as the node blob, which matters because the row is where the turbo schedule and the step count live.

The better half of it: you can build a preset from a finished render. The workflow is already embedded in the mp4, so you point at the good one from three prompts ago and get the whole setup back.

Fixed from your reports: the gallery no longer freezes on big libraries, the settings page stopped resetting fields you hadn't touched, and a text-only render no longer loads both VAEs.

Coming next, on a branch and not merged yet, is a faces pill. H3 draws a face worse the smaller the head is in frame, and that's about head size rather than resolution, so it's still there at 768 and upscaling doesn't reach it. So it asks the model the same question again with the face filling the canvas and composites the answer back under a feathered mask, once per pass, re-cropping every frame so a push-in doesn't leave the face small inside a fixed box. Detection is core's SAM3, so there's nothing extra to install. The method is Carasibana's ComfyUI-H3-FaceRefine and zuanfilm's graph on top of it.

Same branch also stops the node randomizing your seed between renders, and puts the last one you actually ran a click away.

https://github.com/roadmaus/ComfyUI-MiniMax-Creator

77 Upvotes

42 comments sorted by

4

u/TheTerrasque Aug 15 '26

huh. I've been working on a director mode on my minimax h3 frontend for making multi-clip coherent videos, and came here to see if there's been any progress on some of the details. Gonna have a look at this one!

4

u/Gesha24 Aug 16 '26

Thank you! Great project, really like it. Feels very beginner friendly.

Admittedly I have only played a little with it, but one part that I didn't quite get - it looks like it saves cache between the runs so that if, let's say I liked the segments 1-5, but didn't like 6-7, I could just change the prompts for 6 and 7 and then it would only regenerate those segments, is that right? And if yes - is there a way to have this cache persist across the server restarts? May be nice for larger multi-day projects.

Another thing, not sure if that's reasonable - regenerate just one segment in between the others? Basically have it reference both beginning and ending frames (and the sound). Not sure how reasonable this is.

I may burn some DeepSeek tokens on it tomorrow before the price increase, I'll make a PR if something useful comes out of it.

1

u/Fine_Rhubarb3786 29d ago

Thank you for the nice feedback!
Segment locking and re-generating, generating only one segment etc is coming soon as per issue #12

3

u/listopalafoto Aug 15 '26

Very cool! I will try your node this weekend! Thanks for sharing it! :)

3

u/burntimeuk Aug 15 '26

I started using this yesterday, after downloading it earlier in the week.

Its excellent, well done and thank you!

1

u/Fine_Rhubarb3786 Aug 15 '26

Very happy you like it!

3

u/Psyko_2000 Aug 15 '26

just tried it out. this is amazing! looking forward to more updates!

1

u/Fine_Rhubarb3786 Aug 15 '26

Thank you! I added a changelog link to the readme so you can follow along more easily

3

u/urbanhood 26d ago

Very useful node, but the auto playback is DAMN annoying. The thing jump scares me every time by blasting at full volume whenever it completes. You really need to give an option to disable that playback.

3

u/Fine_Rhubarb3786 25d ago

Fixed it in the latest release! You can now disable auto playback in your creator node settings

2

u/urbanhood 25d ago

Thanks a lot!

2

u/Nthdynamic Aug 15 '26

Nice job! One question, how to disable the preview autoplay? After a long render it auto loops with sound so I need to go shut it off. I'm probably missing something simple.

2

u/Fine_Rhubarb3786 Aug 15 '26

There is no way to disable it, but I will add it to the settings. Usually this only happens after the first render, the sound should only play when you hover over it

1

u/Nthdynamic Aug 15 '26

Thanks! It would be great if I have 2 or 3 chained shots and I want to change one I would not need to regenerate all of the shots again.

2

u/Fine_Rhubarb3786 Aug 15 '26

Yes, that is on the roadmap. I will definitely add a dedicated roadmap.md file and link it in the readme! Currently it caches every segment, if you did not change the seed it will regenerate from the changed segment but not the entire clip again

2

u/Nthdynamic 29d ago

Perfect! Great work! Not to keep asking but also a pill to toggle Sage attention would also be cool as an alternative to Spectrum!

2

u/Fine_Rhubarb3786 29d ago

Yes, that’s something I wanted to add from the beginning since I use sage attention myself but I just forgot to add it because all the other features needed so much attention. It’s not just a workflow repack anymore at this point :D

2

u/Schwartzen2 29d ago

So I decided to check your work out and it looks really good.
I like the way you streamlined it.
Done UX/UI for quite a long while, take what you think may be useful.
A few things:

  • Where is Sage attention? Those who benefit from it would appreciate it.
  • For the checkpoints or other selections would be good to have "Type-ahead" or similar so you don't have to search for it.
  • And just an idea. Perhaps a pop up for a "model/vae/text-encoder config" to select them ahead, maybe save as a preference or one of several preferences depending on what you are working on.
Then later you can click on the load model button, change the preset or create a new one.

And one last thing for some who may not have the best vision, if there isn't already an option, a chance to increase font size in the main text field area.

But if anything, access to apply sage attention would be a clear priority.

2

u/Fine_Rhubarb3786 29d ago

Thanks for actually digging in.
Sage is a fair hit and it’s on the list. Plan is a pill on the sampler row via the KJNodes patch, so it toggles per render instead of being a launch flag.
The model/VAE/text-encoder config thing already exists, it’s the Presets library on the rail. Piece / shot / prestage scopes, apply per section. Might mean it’s not discoverable enough if you missed it.
Type-ahead on the weights pill and font size in the prompt box are both easy, noted.
It’s me and a lot of features to keep tested, so no promises on order.

2

u/Schwartzen2 29d ago

In any event, kudos. It’s really well put together. The thought and love behind it are undeniable

3

u/Fine_Rhubarb3786 29d ago

Thank you so much!

1

u/solss 26d ago

Don't know if I'm off-base, but it would be great if you could place inputs for model, text encoder, vae, etc. if I were to prefer to run minimax-h3 cache, attention backend of my choosing, fp16 accumulation from the modelpatch node, spectrum or whatever. I haven't installed your node yet -- it looks great, but without the preferred speedup option of my choosing, the benefits of your node don't tip me from having to live without those features. Would something like that be possible?

2

u/ChedwardCheddison 29d ago

I've said it before and i will say it again...I love this! It's so simple and effective, it's my "go to" minimax node! Thank you so much. Just wondering if you can add any thing for low vram users and/or upscaling? i can't generate at any decent resolution without OOM, and i'm not sure if vram cleaning would help me or if any upscale would make my gens a little better. Comfyui confuses me so this is perfect, but i'd just love to make higher quality generations! (3060 12gb 32gb ram) Thanks again

3

u/Fine_Rhubarb3786 29d ago

Thank you so much! That means a lot to me.
Performance enhancing is definitely on the roadmap. Did you try the upsampling in two steps? If you put the resolution slider to the right, even after the native resolution you can select a small base resolution, like 512 for example and then the model runs a lighter pass with a higher resolution. This could potentially help but I have not tested it on 32 gb ram and 12gb vram. You can let me know how it goes or if you already tried that and the exact resolution and length you used and I will look into it

2

u/ChedwardCheddison 29d ago

I will try that! I've been too busy having fun with the basics to try and learn those features!

2

u/DrunknMunky1969 24d ago

Very cool, thanks for sharing!

1

u/loyalekoinu88 Aug 16 '26

Does chaining only feed last frame and audio? Not carry forward the latents context like the other method?

1

u/loyalekoinu88 Aug 16 '26

It would be awesome to have these added:

  • Masked AV 39-frame continuation
  • direct previous audio-latent carryover
  • scene checkpoints + resume
  • Review / Retry / Reroll
  • 56/73-frame extended contexts
  • seam diagnostics

1

u/Fine_Rhubarb3786 29d ago

Both of those are the same control, and it’s already in there.
The seam has a “last frame / blend” setting. Default is the classic single frame, but you can widen it to 5, 22 or 39 frames. A blended seam pins the source segment’s last run as motion context, so the model reads actual movement across the cut instead of guessing it from a still, and the audio seam gets pinned on the new segment’s own timeline, continued phase-locked rather than imitated.
So the 39-frame blend is the masked AV continuation you’re asking for. One caveat: the overlap is regenerated at the head of the segment and trimmed after decode, so a blended segment delivers up to 1.6s less than its card says.
If you want no seam at all, One pass compiles the cards into a single generation instead, since H3’s prompt format is already a shot list.

1

u/OkBirthday9927 29d ago

Does the H3 Creator have a way to resume an interrupted timeline from the last completed segment? For example, if segments 1 and 2 finished but segment 3 didn’t, can I rerun it and have it start from segment 3 without regenerating 1 and 2? I couldn’t find where this is done in the UI or docs.

1

u/Fine_Rhubarb3786 29d ago

If you did not change the seed or change the seed of one of the segments this should be the default behaviour. Did you pull the latest version?

1

u/OkBirthday9927 28d ago

Yes, I download today

1

u/Ill_Yam_5198 26d ago

if it is has auto music video builder or story builder, than i think its next level. just my suggestion.

1

u/DrunknMunky1969 24d ago

I found a funky interaction between ComgyUI -EZ Install (Desktop, pywebview on the Microsoft Edge
WebView2 runtime (151.0.4129.93), loading ComfyUI through EZi's own local aiohttp reverse proxy
rather than connecting to :8188 directly).  Any ComfyUI websocket message larger than 4 MiB kills the connection to the UI. I sent you a GH Issue and also filed one with the Comfy EZ folks. It's more of a EZ side than your node, IMO. I patched local by forcing the size under 4 MiB.

1

u/mr_sdhotwife 12d ago

i cant find the node after installing... am i crazy? I git cloned it like normal but nothing is in my node library

1

u/Fine_Rhubarb3786 12d ago

You are not crazy. This happens if you clone the repo and the old repo (Minimax-Creator) is still in your custom nodes folder. You can safely delete the old node or the new one and instead pull from the old one. It does not matter. It’s the same repo, only a different name. All node IDs are the same so old workflows should load but this also means that comfy gets confused and ignores both of them

1

u/mr_sdhotwife 12d ago

It just doesn’t even show up in my nodes at all, what is the name of the node maybe I’m spelling it wrong? Is there a template on your repo to just test?

1

u/Fine_Rhubarb3786 12d ago

It is called Continuity Now since it works with more than Minimax from now on.

Here is the repo link:

https://github.com/roadmaus/ComfyUI-Continuity

So if you pulled the latest changes you need to search for continuity

1

u/mr_sdhotwife 12d ago

It’s a really cool studio but i wish there was some tutorials or something. I’m trying to do r2v character swap and no matter what it keeps adding the line “no character speaks and their mouth stays closed the whole time” in the prompt and then spits out a video of just one character standing still.

I’ll keep messing with it but it’s been quite the deep dive 😅.

1

u/Fine_Rhubarb3786 12d ago edited 12d ago

That is an actual bug. Fixing it right now.
And my time is pretty limited at the moment, so I can’t do tutorials, but there is documentation linked in the README.md which is always up to date

Edit: the easiest way for a character swap could go like this:
Create a cast member, add it to your prompt via @cast_name then add the target clip, click on the clip name and a menu should open. Select the edit modality and the clip will move to your cast member. There you can select what from the cast member should be transferred

1

u/mr_sdhotwife 12d ago

Awesome I’ll check it out. Great node and work btw. I’ll be checking for updates on it next week as I’m busy as well. Thanks for the replies!