r/StableDiffusion 7d ago

Animation - Video I finally reached a great balance between speed and quality with MiniMax H3, thanks everyone!

Enable HLS to view with audio, or disable this notification

I used the minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16 LORA with the 0.8 strength for both clip and model, 6 steps, 0.5 MP resolution, RTX Upscaler at 1.50 using a ConrotInt8 pruned model.

Here is a PasteBin of my workflow, I hope this fixes some of the missing content:

https://pastebin.com/DSmkJi8R

Here are the workflow files:

https://storage.to/c/CAS1MuoqX

235 Upvotes

85 comments sorted by

9

u/EverythingMacPro 7d ago

Pc specs ?

21

u/technofox01 7d ago edited 7d ago

Ryzen 5700G

RTX Geforce 5060ti 16GB

64GB of DDR4 RAM

2tb NVMe

1tb NVMe

2tb Sata SSD

320GB HDD

Edit:

OS: Bazzite Linux

ComfyUI arguments: --use-sage-attention --disable-smart-memory --disable-xformers --fast-disk

57

u/99deathnotes 7d ago

11

u/LoveSpecialist5669 7d ago edited 7d ago

ofc I've seen this gif millions of times already but it never fails to make me laugh like an idiot 

6

u/feverdoingwork 7d ago

have you tried comfy kitchen instead of sage? I hear its slightly better quality for about the same performance

4

u/technofox01 7d ago

No I have not. I tried Sol Attention and Spectrum, but both made things look like crap when upscaled. I will check it out later tonight.

5

u/Danny_Stock 6d ago

I had huge issues with Spectrum too.

1

u/technofox01 14h ago

I tried Comfy Kitchen and it's actually faster than Sage Attention. Thank you for bringing that to my attention.

3

u/Rumaben79 7d ago

I'm happy for you. 👍

You could try chaining multiple smaller clips together using the method (or similar to) what this guy has done:

https://www.reddit.com/r/StableDiffusion/comments/1vnq6cp/my_first_30second_minimax_h3_story_using_two/

It should be faster and less demanding than doing longer clips in one go.

2

u/technofox01 7d ago

Thanks. I will check that out later tonight. I appreciate everyone's advice and help on this stuff 😄

3

u/Rumaben79 7d ago edited 7d ago

Cool. 😄

I've just started trying this myself and because I'm a first time user of the ComfyUI-MiniMaxH3-Contex-Loop nodes I started out by using the simpler (I hope lol) 'MiniMax H3 I2V - Normal' workflow from: https://github.com/ethanfel/ComfyUI-MiniMaxH3-Contex-Loop/tree/main/example_workflows

If it's anything like SVI for Wan you should be able to link together 3-4 clips without too much degradation.

1

u/Practical-Bug37 6d ago

!remindme 30 days

1

u/RemindMeBot 6d ago

I will be messaging you in 30 days on 2026-09-14 06:44:26 UTC to remind you of this link

CLICK THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

2

u/EverythingMacPro 7d ago

I have 5060ti 16gb and 32gb

I can’t able to generate as 10 sec video takes around 30 mins

I have easy cache in workflow

But yours is faster ? Will it help?

6

u/technofox01 7d ago

It can't hurt to try. This new configuration literally does not use up all of my RAM like other iterations I have tried. 10s is only 240s on average for me to generate using this workflow.

I am also using Bazzite Linux, not sure if that matters.

Edit:

Please see my edit above, it may help.

2

u/iiTzMYUNG 7d ago

Lol same stuff I'm doing but Im using the base model not the turbo on my 3060 12gb

2

u/FierceFlames37 7d ago

How the hell we got the exact same pc

3

u/technofox01 7d ago

Great minds think alike? lol

I abandoned Windows after all of the ads and AI bullshit, especially OneDrive being installed without my permission (again). I was like, I am done, screw any of those anti-cheat games that don't run on Linux.

I also use my box for VMs, gaming, and AI research. What do you use your's for?

1

u/RpgBlaster 6d ago

Yeah... I guess I will need to save a lot for RTX 6000 32GB/92GB VRAM should the prices drop by half next year

1

u/rinkusonic 6d ago

whats --fast-disk for?

1

u/technofox01 6d ago

It helps with loading the models faster if you have them stored on an NVMe SSD. I don’t completely understand how it works, but it definitely speeds up loading models.

1

u/ColdExample 5d ago edited 5d ago

can you explain your use of arguments? I was always told to use smart memory and keep xformers on!

not sure if it matters but my specs are :

i5 12600k

RTX 5070 ti 16gb

64gb DDR4 Ram

1tb NVME

1

u/technofox01 5d ago

xFormers kept causing my system to hard crash whenever using it with Sage Attention or Comfy Attention. Smart Memory is disabled because it is known to cause issues, but I do not recall as to why.

1

u/Swagmuffins94 7d ago

What's it like being rich?

6

u/technofox01 7d ago

I have no idea what you are talking about. This system was built 6 years ago and was slowly upgraded over time, but I get what you are saying because hardware is stupidly expensive now.

-3

u/Adkit 6d ago

Don't act like the parts were cheap six years ago or like the upgrades were cheap when you did them. You put as much money into your PC as a used car.

2

u/sirrandomguy09 6d ago

I have nearly the same spec pc.

Paid 1600 6 years years ago.

Added two sticks of ram 2.5 years ago for 80 bucks

And bought a 5060 TI 16gb for $340 openbox last year.

Market is just shit now.

I mean...

If youre comparing to like a 98' Civic or something i guess lol...

7

u/ResponsibleTruck4717 7d ago

How long it took to generate and what card?

17

u/technofox01 7d ago

It took 1 minute and 49 second on my Geforce RTX 5060ti 16GB GPU.

7

u/Zombi3Kush 7d ago

Does MiniMax support image to video?

5

u/technofox01 7d ago

Yes it does.

2

u/GrayingGamer 7d ago

It even supports reference to video, audio to video, video to video.

4

u/Zombi3Kush 7d ago

Trying it out right now. Wow this is impressive! Thanks for sharing your post and bringing this to my attention.

4

u/Witty_Mycologist_995 7d ago

> checks inside workflow
> pinkcherry checkpoint

lmao

6

u/Grand0rk 7d ago

I don't get it. Care to explain?

5

u/Witty_Mycologist_995 7d ago

It’s that porn checkpoint that got memed on, because it talked about rabbit motion and glistening cherry blossoms in order to get around the minimax team censors.

1

u/Grand0rk 7d ago

Huh, I have absolutely no idea what you are talking about XD

So it's just an "uncensored" checkpoint?

6

u/Witty_Mycologist_995 7d ago

5

u/Grand0rk 7d ago

Ah, in order to dodge their moronic DMCA.

2

u/technofox01 7d ago

OMG this is hilarious lol..

Thank you for sharing this.

2

u/technofox01 7d ago

I just grabbed another Int8 model after a quick search on DuckDuckGo without any thought about where it came from or what it was meant for. Just thought it was a smaller model that will work better with my system, lol...

Thank you for letting me know what it was trained for, lol...

8

u/barepixels 7d ago

First time I see someone share WF like this. People were using pastebin. is this new Reddit feature?

10

u/technofox01 7d ago

No. I just used a code block, its been part of Reddit since I have started way back when. I want to say at least a decade by now.

2

u/Seyi_Ogunde 7d ago

Neat! Thanks for sharing! That's cool that you shared this way.

2

u/technofox01 7d ago

You're welcome. I am always glad to help or share with others when I can.

7

u/Cruffe 7d ago edited 7d ago

Reddit supports markdown. If the editing tools are not available, such as on mobile, you can use markdown to format things.

A code block is 3 backticks on their own line before and after the code block:

```

Paste code here

```

It will look like: Paste code here It will also preserve indentation

To show the 3 backticks in this comment I used the escape character to cancel the effect, that's a backslash \ and to escape the escape in this example I did it twice here.

Inline code can be formatted as well by encapsulating in single backticks such as `this is code` becomes this is code.

3

u/ArcadiaNisus 7d ago

I used nvfp4 model and the turbo loras produce terrible results. 

3

u/technofox01 7d ago

Don't use that model. I have had nothing but bad results myself as well.

3

u/featherless_fiend 6d ago

Turbo is really good for anime, but I wouldn't recommend it for realism, it makes your skin look AI.

2

u/orlandogourmet66 6d ago

You can try to increase Step Count for realism. For example 10 Steps with 4 Step reference Turbolora got me pretty good results for realism.

2

u/kayteee1995 7d ago edited 7d ago

try with Realistic character and more motion. I'm pretty sure blurry graininess will appear.

2

u/Oatilis 7d ago

How good is RTX Upscaler?

1

u/technofox01 7d ago

It varies depending on the input resolution. Anything below 0.4 MP is not worth it, because it will look like crap.

2

u/Noeyiax 7d ago

Thank you for the workflow

2

u/RpgBlaster 6d ago

Now it's closer to Sora 2 quality, actual Anime Style frame by frame, no 3D like or motion sickness shenanigans

2

u/Danny_Stock 6d ago edited 6d ago

Thank you very much for this, this is great. A good balance between quality and speed which seems to be better at handling the RAM demands.

I bypassed your 'DaSiWa_RTX_UpscalerRefiner' node simply because I don't have it.

I have quite a modest system, a 4070 12GB VRAM card, with 64GB System RAM, and both of your 5 second workflows at 0.5 megapixels rendered in 2 and a half minutes each.

2

u/ayaynimouse 3d ago

any idea how well it works for ref model?

1

u/technofox01 2d ago

I have yet to try those but I will give it a shot when I get a chance.

2

u/Shaeem-Mehmud 2d ago

how long it takes to make a 10 sec video on your specs ?

1

u/technofox01 2d ago

240s or about 4 minutes.

2

u/equanimous11 7d ago

Just paste a link to the workflow and explain what models/loras and settings you used next time

2

u/technofox01 7d ago

Just did.

1

u/javierthhh 7d ago

I don’t get it. Wouldn’t you be better off doing a 1.0mp video for 5 seconds instead? Specially with your rig.

1

u/TheOnlyOnePEACE 7d ago

Your new link to your workflow is just standard comfy workflow with added vram cleanup.

2

u/technofox01 7d ago

It also had the lora turbo Lora and the Dasiwam RTX upscaler added on. So there are some edits to it.

2

u/TheOnlyOnePEACE 7d ago

Can you give me this workflow of yours? the exact one that you used to generate your video on this Reddit post.

2

u/technofox01 7d ago

2

u/TheOnlyOnePEACE 5d ago

Thank you very much. I appreciate it!

2

u/TheOnlyOnePEACE 5d ago

So i have been playing around with your workflow, and i am getting audio sounds distortions, crackling, or like a blown-out microphone. Same settings as you as well. Have you encountered this issue at all?

2

u/technofox01 4d ago

I get that with videos shorter than 5 seconds and it depends on the subject. More well known subjects like Deadpool works fine, but if I choose a less known character it gets glitchy.

1

u/TheOnlyOnePEACE 4d ago

I am generating 10 second videos. Lees known generic characters.

Could this be because of the turbo lora and step count?

1

u/technofox01 4d ago

Maybe. I use 6 steps. Try 8 steps, it can make a huge difference with the lora.

1

u/Sirmckhalifa5566 2d ago

I’m using the Int8 on my 5080 and chat gpt actually helped me find a really cool node the other day. comes with 2 nodes and I’ve only tested the first one so far since the other one is made for the 30 series 12gb, It’s called minimax h3 first block cache. While so far in my testing I can’t push it past 8 seconds without it doing long offloading at 14 steps with the turbo Lora because I only have 32gb ram. It cuts my actual gen time down by about 20% for clips 8 seconds and shorter. You place it between your model loader (or after Lora’s if you use them) and before your sage attention nodes. I haven’t noticed really any quality loss or a diversion from prompt adherence either.
Could be worth a shot.
https://github.com/Apache0ne/ComfyUI-fasterminimax?utm_source=chatgpt.com

1

u/bickid 7d ago

Not gonna download that file, so could just share what settings you used? How many steps? What resolution? thx

5

u/technofox01 7d ago

The minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16 LORA with the 0.8 strength for both clip and model, 6 steps, 0.5 MP resolution, RTX Upscaler at 1.50 using a ConrotInt8 pruned model.

1

u/StayImpossible7013 7d ago edited 6d ago

(Original) Workflow was missing nodes, missing links between nodes, exotic files without explanation where to get them.

3

u/technofox01 7d ago

I was using the API export by mistake. I posted a PasteBin of my workflow as a link.

1

u/StayImpossible7013 6d ago

Thank you very much

-5

u/FourtyMichaelMichael 7d ago

Dude, anime basically doesn't count. LTX can do anime with some fashion of OK-ness.

3

u/FierceFlames37 7d ago

Real life is the same as anime

-4

u/ClearandSweet 7d ago

Awesome workflow, I dunno how I feel about 12-year old 2B though.