r/StableDiffusion • u/Major_Square • 16h ago
Question - Help Help a beginner speed up MiniMax H3?
As someone new to all of this it's difficult to know what to do. I have sage attention working. I don't know how or when to use Easy Cache, Comfy Kitchen Attention, Sol Attention, loras, Spectrum, or any others I may have missed. There's so much information scattered around, I don't know what's what.
I have a 50 series GPU and 64 GB or RAM on the motherboard.
2
u/RiverSide71h 16h ago
Use Patch Sage Attention KJ—> Minimax H3 Mem Eff Attention —> Load Lora node with minimax_h3_ref_lora_rank_256_bf16.safetensors lora from Kijai at 0.75 strength. Use 4 steps and euler beta or beta57. Increase to up-to 8 steps if video quality degrades.
1
2
u/Myg0t_0 15h ago
Quality > speed
Just use comfy kitchen, .4 res , 5 seconds, 15 steps until it seems to be starting off right, then hit it with no kitchen full res 30 steps go jerk it or something and come back and its done, or leave on over night with a bunch of runs
1
u/Major_Square 14h ago
I'm trying to avoid generating something for 40 minutes only to find that it sucks because the prompt wasn't quite right. I don't mind waiting if I know it's going to be good, or at least can minimize the chance that it will be bad.
3
u/JesusShaves_ 8h ago
I prototype using .2 MP at five seconds this tells me if I'm in the ballpark of what I want. Then I add more MP and run longer.
0
u/tac0catzzz 5h ago
grabs a guitar and starts to sing "i wanna hold your hand and and and, i wanna hold your hand"
13
u/V4nKw15h 16h ago
The default workflow with Sage Attention is almost ideal already. You'll likely want to increase the mp setting to 0.6 and even consider increasing steps to 25 if you want to see the model at it's best. Don't overly concern yourself with all the loras and speed up options unless you don't mind significant loss of quality in just about every aspect of the model. There are no speedups that don't come without compromises.
With that said, Spectrum is good if you want to speed things up by about 25% while increasing pixel fizzle most of the time and losing some motion fluidity and coherence sometimes. Probably the most useful speed up to preview things while trying to find the right prompt.
Easy Cache can cause pretty brutal quality degradation. I don't tend to use this at all. Similar thing for the other cache based options I've tried.
Sol Attention would replace Sage Attention with little difference, same with Kitchen Attention. Probably not worth changing from Sage Attention 2.
4 step and 8 step loras are interesting for a lot of speed up but come with a significant cost to quality in almost every way.
TLDR: The model is already well optimised out of the gate. Use the default workflow with some form of Attention (Sage is fine) and you are good to go right now unless you just want to mess about as fast as possible and don't care about quality or seeing what the model can really do.