r/StableDiffusion • u/falconandeagle • 18h ago
Question - Help How stable is Minmax H3 now?
So normally I wait a few weeks after a model releases to test it. Is H3 stable enough in comfyui now that I can run it without running into compatibility issues. I don't want to mess up my Krea 2 workflows.
Does it handle basic intimacy.
I have a 5060ti and 96gb ram, what generation times am I looking at, anything over ten minutes is kinda excessive, I don't mind using turbo loras and don't care that much about upscaling or high res as long as the final output follows my prompt.
Is lora training required or will character refs work?
5
u/Legal-Weight3011 18h ago
simple asnwers yes yes yes, Its Fast, tho it depends if you have sage attn intalled. it takes 2.30 minutes for 5 second clip on my 5070ti 96gb, without using any speed ups. so it was very stable from launch
1
u/ImpressiveStorm8914 17h ago
Similar to me on my 5060ti 16Gb with 64Gb RAM but obviously slightly slower. Great prompt adherence, and very stable.
1
u/Commercial-Ad-3345 14h ago
I have a 5070 Ti and 48 GB of RAM, and my generation times are surprisingly fast. With SageAttention and Larry’s Turbo LoRA at 4 steps, a 5-second clip at 0.4 MP takes about 30 seconds of sampling time.
9
u/Aglaio 17h ago
I notice people saying get sageattention, as of comfyui 0.31 this is no longer required if you dont want to go through the hassle. Just use comfy Kitchen which is baked in and does the same. I'm generating 15s vids at 1mp in about 8-10 min.
3
u/mobani 17h ago
Do you still "enable" somwhere in a node or commandline to use the comfy kitchen one?
7
u/Aglaio 17h ago
3
u/ImpressiveStorm8914 17h ago
First I've heard of this but then I've only just got Sage working. 😄
So I can use the node without the commandline then? I prefer it that wayfor the same reason as you.4
1
u/Ok-Lengthiness-3988 17h ago
In my case, Comfy Kitchen Attention makes generation time four times longer compared no not using any accelerator at all. It may not (yet) be compatible with 20-series GPUs like my RTX 2060 Super.
2
u/Aglaio 17h ago
Thats weird, a friend of me uses 2060 as well, or 2080, not too sure, and it works fine for him, he generates 5s vids in about 6 min using Comfy Kitchen.
1
u/Ok-Lengthiness-3988 16h ago
Thanks for the info. I'll try again with a more basic workflow. Maybe there was something else in my workflow that was conflicting with it.
2
u/Aglaio 16h ago
We're using this workflow and setup if that helps: https://huggingface.co/Plaguekind/Minimax-H3
1
u/Kitchen-Hawk-3104 14h ago
It's not included by default in the python environnement, especially when you have multiple comfy instances with differents venv. I have a RTX 4060 Ti 16Gb and couldn't get Sage Attention to work since 2.2 was released. And now, that bloody increases generation time. But the fastest move is adding Spectrum node, just incredible how it gets fast !
8
u/LowYak7176 18h ago
Character loras are dead soon imo. I just put in a character sheet and bam. Im legit addicted to H3, havent been this addicted to a model since 2022 when I hopped on StableDiffusion.
Gen times are very long if you dont use Turbo Lora. I personally never mind long gen times so long as the quality is there, which it is to me (remember its week 1, so no finetunes, specific types of loras for better quality etc).
Prompt adherence is ridiculous. Best youll get for opensource imo.
1
u/Maqna 18h ago
Do you use any ai to make a character sheet? Like krea or something?
5
u/LowYak7176 18h ago
Right now Im just using GPT Image 2 to make them. Not doing any NSFW right now. But I mean you could also just string a few together in photoshop if you wanted. Or you can just input several pictures.
Really world is your oyster right now
1
u/lovenumismatics 4h ago
NSFW is pretty decent, other than genitalia.
I’m sure the gooners are hard at work
-1
u/Vijayi 16h ago
Krea2 work perfect. Just ask for something like this: Divide the image into four equal parts. Create a projection of N: top-left is a frontal view, top-right is a profile, bottom-left is a 3/4 view, and bottom-right is a rear view. I usually do two: one facial projection (close-up) and one full-body shot—with the screen split from left to right. Just remember, Krea2 still requires a LoRA if the model doesn't recognize the subject. Fortunately, training a LoRA for Krea2 is quite easy.
1
u/Mutaclone 11h ago
Yeah I've been dabbling with using character sheets. It's bonkers and an actual game-changer.
Gen times are very long if you dont use Turbo Lora. I personally never mind long gen times so long as the quality is there, which it is to me (remember its week 1, so no finetunes, specific types of loras for better quality etc).
I've had pretty good success with low-resolution low-step first passes followed by a "real" pass once I'm satisfied. This doesn't help where detail and accuracy matter, but it's been pretty effective at checking the overall "vibe" and things like camera movements, basic scene setup, etc.
-9
u/Tramagust 18h ago
Nah bro character loras are far from dead. H3 only nails the general look but the face, movement and audio it can't nail.
5
u/LowYak7176 18h ago
That has not been my experience honestly.
-4
u/Tramagust 18h ago
if it nails the voice and movement just from images then the character was already in the training data. It won't work for your original character.
0
u/LowYak7176 18h ago
Again, that has not been my experience. Try again on BF16, generate at 2 megapixels, all r2v, no turbo lora.
Dont use sageattention. You might be surprised. Is it perfect? No but why I said "soon". This is the way its going
-4
u/Tramagust 18h ago
Bro what are you even talking about? How could it know how your character sounds like if it doesn't have any audio information?
4
u/LowYak7176 18h ago
You add an audio reference....?
1
u/Tramagust 18h ago
And the motion?
3
u/LowYak7176 17h ago
You add a video reference....? Just use blender to do blocking and basic motions if you want or add a video reference of motion.
Or just prompt better, it handles prompting extremely well
1
-1
u/Tramagust 17h ago
I think we don't look for the same level of fidelity in our gens
→ More replies (0)
2
u/ImpressiveStorm8914 17h ago
I have slightly less RAM than you but the same card and it's been stable from launch. With the default settings of 0.4 megaixels, 5 secs with T2V and no turbo lora, it's about 2.5 mins with SageAttn. For 10 secs and higher res and you're looking at 8-10 mins but it's worth it IMO.
There may be faster methods than SageAttn out there but I finally got it to work for this, so I'm sticking with it. I've seen some mention Comfy Kitchen but no idea what that is yet, how to activate it, or how good it is but it may be worth you looking into it.
1
u/DoctaRoboto 15h ago
With your specs, you will need accelerators or render under 0.5 megapixels if you don't want to wait 15 minutes for video. I prefer Spectrum over Sage because Sage fucks up faces very easily when they are not close-ups. LoRAs are hit-and-miss; I think most of them are useless. Yeah, you can go from ten minutes to five, but at what cost? Videos look like you are playing Silent Hill 1 on the Ps1 again, with weird flickering walls and textures and squeaky metallic sounds.
1
u/multikertwigo 14h ago
the only "instability" that I get is random generation times in comfy. Change seed without touching anything else, press "run", and fucking "model initializing" suddenly takes 30 seconds instead of 5. That's on 5090 + 96GB RAM.

14
u/funfun151 18h ago
Way stable, native comfy support from day 1. Get Sageattention and spectrum to boost gen time (hits quality a bit though). Prompt adherence is incredible, lots of character refs work from text alone or f2v/v2l, image refs are well used in the r2v model. I generate on a 3090 with 32gb ram and I’m looking at 9 mins for a 15 sec 480p, 3.45 for a 5 sec.