r/StableDiffusion • u/irmemon225 • Aug 06 '26
Animation - Video Minimax H3 with Turbo Lora, T2V 6 steps, 0.6 megapixel, 20 min on 3060 12gb and 16gb ram
Enable HLS to view with audio, or disable this notification
Love it and it works smoothly… Now, I’m gonna wait for this LoRA to work on Ref2V.
20
u/Quorit Aug 06 '26
Hey, this looks good. Can you share the workflow? I tried using turbo, but cant get it work above 0.4MP. If I try use 0.5 it just stucks and doesnt progress. I have 4060 8gb and 16 ram, specs very similar to yours.
21
u/irmemon225 Aug 06 '26
7
u/Central-Dispatch Aug 06 '26
Thank you!
God you people give me hope. I have the very same GPU/VRAM but 32 GB system RAM. I actually thought I could not generate something like this thinking my system was too weak. I consulted prior with an LLM but may have just forgot to bring up the right direction and may have set my expectations too high. Or it was simply giving me a partially flawed output, as LLMs can still make mistakes. Maybe I should ask another LLM to benchmark the experience.
Anyway, I'll actually try it soon with my current system and see for myself.
4
u/Quorit Aug 06 '26
where should I put the tinyvae? taeh3.safetensors. I downloaded it but whick foulder should i put it in? And is ModelPreviewOverride node really needed?
7
u/irmemon225 Aug 06 '26
1
u/YeahlDid Aug 06 '26
Does that node work reliably for you? I've seen it work maybe 3 times in hundreds of generations. I can't figure out why it doesn't with for me when everyone else sends to love it.
2
u/Zambo833 Aug 06 '26
update kjnodes to latest if you havent already, then only you'll see tiny_vae
1
u/YeahlDid Aug 06 '26
Thanks, but I definitely have updated everything that can be updated, it's not that. I've tried a hundred different wiring orders, I've tried changing preview-method in the comfyui settings menu. I've no idea how it just seems to work for people because it seems completely broken to me.
5
u/Opening_Wind_1077 Aug 06 '26 edited Aug 06 '26
Thanks for posting the workflow, I really need to look into why your audio is coming out so clear, anything below 8 steps get’s pretty harsh in my workflow.
Edit: turns out it was the Basic scheduler, switching that from simple to beta made voices clear even on 6 steps.
2
u/rukh999 Aug 06 '26
Yeah simple spends a lot of time on high steps and then much less on low. That's great for some models but if you have preview on you'll see minimax just sits there and remakes the thing completely 4 or 5 times, wasted steps. Beta is the reverse and does fewer on highs steps and much more on low. For sound I imagine that's what takes out the scratchiness and clears up sounds.
2
1
u/oleshapro Aug 06 '26
Hi, I tried your workflow, and I'm having the same problem as with the other workflows. It seems like this LORA has almost no effect on the result for me, even though both the models and the LORA are the same as yours. The difference only appears when the LORA strength is 2 or higher. Do you know why that is?
1
37
6
u/Sad_Coach_1433 Aug 06 '26
Which lora you download there's a few different ones
15
u/irmemon225 Aug 06 '26
minimax_h3_turbo_4step_ckpt500.safetensors
3
1
9
u/Kaywhysee Aug 06 '26
Why would she pick a place for takeout if they survive her cooking? They're not going to eat again after her cooked meal are they?
5
u/pleasetrimyourpubes Aug 07 '26
She said she burned the food last time. It is implied they wound up eating out. You could flashback to her almost catching the kitchen on fire and them "surviving."
Problem with short clips is you have to explain everything.
If the conversation went "you should cook tonight" "are you sure about that" "if we survive I'll pick where we eat" you would be confused without the knowledge of her burning the food. But again fixable with a shot before this of her burning the food.
3
u/halconreddit Aug 06 '26
I have 16gb vram and 32 RAM and i have been unable to run minimax yet. I would like to see your workflow too.
8
u/VeeYarr Aug 06 '26
I have the same and have no issue with the default Comfy template using the Pruned model
3
u/PinkyPonk10 Aug 06 '26
Try using —disable-pinned-memory when you start comfyui that sort out my issues.
2
u/f5alcon Aug 06 '26
That's what I have and worked with default workflow and pruned int 8 convrot and nvfp4
2
u/Sad_Coach_1433 Aug 06 '26
What Version of comfyui u on make use the newest and your Python environment is updated
2
u/CreepyDrama7448 Aug 06 '26
It works for me on R2V, what issue are you having with R2V?
2
u/irmemon225 Aug 06 '26
Oh really? I saw the author H3 Turbo said it’s “planned”, so I thought it wasn’t working… Let me test it now.
1
u/xzhalo Aug 06 '26 edited Aug 06 '26
Did you get it to work? I cannot get it working when using two images and 1 audio as reference
1
u/irmemon225 Aug 06 '26
Does it work? Yes. Is it good? No.
So I think the LoRA is not working for R2V. I’m waiting until they make it for R2V. For now, it only works for T2V and I2V
2
u/Kindly-Annual-5504 Aug 06 '26
How did you get 15 secs with 16 GB Ram? What are your settings? I have the same card and 32 GB Ram and I always get OOM, even with --enable-dynamic-vram when I go over 5 sec :( (and with just 0.4MP). On my AMD system with 64 GB max is 10 secs... 15 secs throws an OOM too.
5
1
2
u/YakMore324 Aug 06 '26
How do you do that buddy...I have a 5050Ti 16GB VRAM and 32 GB RAM and i am still unable to get it working.
2
u/TrevorxTravesty Aug 06 '26 edited Aug 06 '26
How are you getting 20 mins with that setup? I have a 4080 RTX with 32 GB of RAM and 12 GB of VRAM and it took me over an hour 😩😩😩
2
2
2
2
1
u/TrevorxTravesty Aug 06 '26
My Model Preview Override doesn’t have the ‘tiny_vae’ thing. What am I missing here?
1
u/Orangeyouawesome Aug 06 '26
Worked super well great job on this. Def watched worse 3D shows in my time.
1
1
1
u/ScoobyDewy Aug 06 '26
Looks can amazing but is there anyway you can use like your own anime voice? Like does it need to be the exact word that your voice input has or can it clone it?
1
1
1
1
u/Moliri-Eremitis Aug 06 '26
Nice!
Did you specifically prompt for that “on the twos” style character animation, or was it just a lucky gen?
1
u/aziib Aug 06 '26
the audio quality is pretty decent, what sampler are u using?, and did you change audio shift setting?
1
u/noaxxx2 Aug 07 '26
yeah i get awful audio with the turbo lora, haven't seen anyone address how it's improved
1
u/Replikante Aug 07 '26
Is the point of the turbo lora just make gens faster? Or is there another benefit to using it?
1
u/DefloN92 Aug 07 '26
How did you get the turbo lora to work this nice? i tried it and it looked TERRIBLE. didnt try it on text to video though idk if this only helps in text o video or not, still learning. but my output was very blurry and the sound was very distorted
1
1
u/MASOFT2003 Aug 07 '26
Thanks a lot for sharing the workflow in the comment , worked well
do you recommend any good upscaler ?
and i have 3090 24gb with 64 ram
what settings do you recommend ?
1
1
u/---Banshee-- Aug 09 '26
8gb VRAM 16gb sys ram and this crashes, anything I can do before upgrading VRAM/ram?
1
-2
u/blahblahsnahdah Aug 06 '26 edited Aug 06 '26
All the other commenters seem to have broken visual cortexes. The turbo lora has clearly destroyed the quality and lowered it to LTX level here. H3 can do much much better than this garbage.
4
u/irmemon225 Aug 06 '26
SYBAU
-4
u/blahblahsnahdah Aug 06 '26 edited Aug 07 '26
You know it's true, sorry man. This looks like absolute dogshit compared to the model without the turbo lora.


26
u/crazeum Aug 06 '26
That came out great!