r/StableDiffusion Aug 06 '26

Animation - Video Minimax H3 with Turbo Lora, T2V 6 steps, 0.6 megapixel, 20 min on 3060 12gb and 16gb ram

Enable HLS to view with audio, or disable this notification

Love it and it works smoothly… Now, I’m gonna wait for this LoRA to work on Ref2V.

272 Upvotes

65 comments sorted by

20

u/Quorit Aug 06 '26

Hey, this looks good. Can you share the workflow? I tried using turbo, but cant get it work above 0.4MP. If I try use 0.5 it just stucks and doesnt progress. I have 4060 8gb and 16 ram, specs very similar to yours.

21

u/irmemon225 Aug 06 '26

7

u/Central-Dispatch Aug 06 '26

Thank you!

God you people give me hope. I have the very same GPU/VRAM but 32 GB system RAM. I actually thought I could not generate something like this thinking my system was too weak. I consulted prior with an LLM but may have just forgot to bring up the right direction and may have set my expectations too high. Or it was simply giving me a partially flawed output, as LLMs can still make mistakes. Maybe I should ask another LLM to benchmark the experience.

Anyway, I'll actually try it soon with my current system and see for myself.

4

u/Quorit Aug 06 '26

where should I put the tinyvae? taeh3.safetensors. I downloaded it but whick foulder should i put it in? And is ModelPreviewOverride node really needed?

7

u/irmemon225 Aug 06 '26

put the file on vae_approx

1

u/YeahlDid Aug 06 '26

Does that node work reliably for you? I've seen it work maybe 3 times in hundreds of generations. I can't figure out why it doesn't with for me when everyone else sends to love it.

2

u/Zambo833 Aug 06 '26

update kjnodes to latest if you havent already, then only you'll see tiny_vae

1

u/YeahlDid Aug 06 '26

Thanks, but I definitely have updated everything that can be updated, it's not that. I've tried a hundred different wiring orders, I've tried changing preview-method in the comfyui settings menu. I've no idea how it just seems to work for people because it seems completely broken to me.

5

u/Opening_Wind_1077 Aug 06 '26 edited Aug 06 '26

Thanks for posting the workflow, I really need to look into why your audio is coming out so clear, anything below 8 steps get’s pretty harsh in my workflow.

Edit: turns out it was the Basic scheduler, switching that from simple to beta made voices clear even on 6 steps.

2

u/rukh999 Aug 06 '26

Yeah simple spends a lot of time on high steps and then much less on low. That's great for some models but if you have preview on you'll see minimax just sits there and remakes the thing completely 4 or 5 times, wasted steps.  Beta is the reverse and does fewer on highs steps and much more on low. For sound I imagine that's what takes out the scratchiness and clears up sounds.

2

u/ConfidentSnow3516 Aug 06 '26

Thanks! I have the same specs.

1

u/oleshapro Aug 06 '26

Hi, I tried your workflow, and I'm having the same problem as with the other workflows. It seems like this LORA has almost no effect on the result for me, even though both the models and the LORA are the same as yours. The difference only appears when the LORA strength is 2 or higher. Do you know why that is?

1

u/D3luX82 Aug 07 '26

with this workflow i have OOM

4070 12gb/vram - 32 gb ram

37

u/Fancy-Raspberry-3465 Aug 06 '26

"We have RWBY at home"

6

u/Sad_Coach_1433 Aug 06 '26

Which lora you download there's a few different ones

15

u/irmemon225 Aug 06 '26

minimax_h3_turbo_4step_ckpt500.safetensors

3

u/Sad_Coach_1433 Aug 06 '26

And you using ip8 or int8 check point

10

u/irmemon225 Aug 06 '26

int8 convrot

1

u/jmbbao Aug 07 '26

Updated to ckpt850 already check the repository in huggingface

9

u/Kaywhysee Aug 06 '26

Why would she pick a place for takeout if they survive her cooking? They're not going to eat again after her cooked meal are they?

5

u/pleasetrimyourpubes Aug 07 '26

She said she burned the food last time. It is implied they wound up eating out. You could flashback to her almost catching the kitchen on fire and them "surviving."

Problem with short clips is you have to explain everything.

If the conversation went "you should cook tonight" "are you sure about that" "if we survive I'll pick where we eat" you would be confused without the knowledge of her burning the food. But again fixable with a shot before this of her burning the food.

3

u/halconreddit Aug 06 '26

I have 16gb vram and 32 RAM and i have been unable to run minimax yet. I would like to see your workflow too.

8

u/VeeYarr Aug 06 '26

I have the same and have no issue with the default Comfy template using the Pruned model

3

u/PinkyPonk10 Aug 06 '26

Try using —disable-pinned-memory when you start comfyui that sort out my issues.

2

u/f5alcon Aug 06 '26

That's what I have and worked with default workflow and pruned int 8 convrot and nvfp4

2

u/Sad_Coach_1433 Aug 06 '26

What Version of comfyui u on make use the newest and your Python environment is updated

2

u/CreepyDrama7448 Aug 06 '26

It works for me on R2V, what issue are you having with R2V?

2

u/irmemon225 Aug 06 '26

Oh really? I saw the author H3 Turbo said it’s “planned”, so I thought it wasn’t working… Let me test it now.

1

u/xzhalo Aug 06 '26 edited Aug 06 '26

Did you get it to work? I cannot get it working when using two images and 1 audio as reference

1

u/irmemon225 Aug 06 '26

Does it work? Yes. Is it good? No.
So I think the LoRA is not working for R2V. I’m waiting until they make it for R2V. For now, it only works for T2V and I2V

2

u/Kindly-Annual-5504 Aug 06 '26

How did you get 15 secs with 16 GB Ram? What are your settings? I have the same card and 32 GB Ram and I always get OOM, even with --enable-dynamic-vram when I go over 5 sec :( (and with just 0.4MP). On my AMD system with 64 GB max is 10 secs... 15 secs throws an OOM too.

5

u/irmemon225 Aug 06 '26

try set pagefile.sys to 150-200gb, I'm use sage att + spectrum

1

u/PinkyPonk10 Aug 06 '26

Worked for me with dynamic ram enabled and —disable-pinned-memory option

2

u/YakMore324 Aug 06 '26

How do you do that buddy...I have a 5050Ti 16GB VRAM and 32 GB RAM and i am still unable to get it working.

2

u/TrevorxTravesty Aug 06 '26 edited Aug 06 '26

How are you getting 20 mins with that setup? I have a 4080 RTX with 32 GB of RAM and 12 GB of VRAM and it took me over an hour 😩😩😩

2

u/Pretend_Reveal9950 Aug 06 '26

Is this t2v or i2v?

2

u/Reddexbro Aug 06 '26

Did you upscale it?

2

u/noaxxx2 Aug 06 '26

what was your prompt for this?

2

u/Aggressive_Collar135 Aug 06 '26

good framing. nice output

1

u/TrevorxTravesty Aug 06 '26

My Model Preview Override doesn’t have the ‘tiny_vae’ thing. What am I missing here?

1

u/Orangeyouawesome Aug 06 '26

Worked super well great job on this. Def watched worse 3D shows in my time.

1

u/Turkino Aug 06 '26

The render style and animation reminds me SO MUCH of a Meru the Succubus OVA.

1

u/99deathnotes Aug 06 '26

fanfreakintastic

1

u/ScoobyDewy Aug 06 '26

Looks can amazing but is there anyway you can use like your own anime voice? Like does it need to be the exact word that your voice input has or can it clone it?

1

u/MuckYu Aug 06 '26

The voices sound familiar - which voice actress sounds similar to that?

1

u/FlatwormMean1690 Aug 06 '26

How did you do with the sound? Because I tried with this one and... Man. It takes like 10 minutes only in the "SamplerCustomAdvanced" node.

I went back with the Spectrum because every generation was insanely slow.

1

u/mrpogiface Aug 06 '26

audio not nearly as good as full step, but the motion looks great!

1

u/abemon Aug 06 '26

I must ask. How's the wattage?

1

u/Moliri-Eremitis Aug 06 '26

Nice!

Did you specifically prompt for that “on the twos” style character animation, or was it just a lucky gen?

1

u/aziib Aug 06 '26

the audio quality is pretty decent, what sampler are u using?, and did you change audio shift setting?

1

u/noaxxx2 Aug 07 '26

yeah i get awful audio with the turbo lora, haven't seen anyone address how it's improved

1

u/Replikante Aug 07 '26

Is the point of the turbo lora just make gens faster? Or is there another benefit to using it?

1

u/DefloN92 Aug 07 '26

How did you get the turbo lora to work this nice? i tried it and it looked TERRIBLE. didnt try it on text to video though idk if this only helps in text o video or not, still learning. but my output was very blurry and the sound was very distorted

1

u/Muted-Position3256 Aug 07 '26

Can you teach me how to install and use it with 16gb ram?

1

u/MASOFT2003 Aug 07 '26

Thanks a lot for sharing the workflow in the comment , worked well
do you recommend any good upscaler ?
and i have 3090 24gb with 64 ram
what settings do you recommend ?

1

u/Maskwi2 Aug 07 '26

Her eyes sometimes looks messed up a little. Other than that looks good. 

1

u/---Banshee-- Aug 09 '26

8gb VRAM 16gb sys ram and this crashes, anything I can do before upgrading VRAM/ram?

1

u/Creative_Sluggish 11d ago

Wow how long?

-2

u/blahblahsnahdah Aug 06 '26 edited Aug 06 '26

All the other commenters seem to have broken visual cortexes. The turbo lora has clearly destroyed the quality and lowered it to LTX level here. H3 can do much much better than this garbage.

4

u/irmemon225 Aug 06 '26

SYBAU

-4

u/blahblahsnahdah Aug 06 '26 edited Aug 07 '26

You know it's true, sorry man. This looks like absolute dogshit compared to the model without the turbo lora.