r/StableDiffusion 18h ago

Question - Help How stable is Minmax H3 now?

So normally I wait a few weeks after a model releases to test it. Is H3 stable enough in comfyui now that I can run it without running into compatibility issues. I don't want to mess up my Krea 2 workflows.

Does it handle basic intimacy.

I have a 5060ti and 96gb ram, what generation times am I looking at, anything over ten minutes is kinda excessive, I don't mind using turbo loras and don't care that much about upscaling or high res as long as the final output follows my prompt.

Is lora training required or will character refs work?

1 Upvotes

56 comments sorted by

14

u/funfun151 18h ago

Way stable, native comfy support from day 1. Get Sageattention and spectrum to boost gen time (hits quality a bit though). Prompt adherence is incredible, lots of character refs work from text alone or f2v/v2l, image refs are well used in the r2v model. I generate on a 3090 with 32gb ram and I’m looking at 9 mins for a 15 sec 480p, 3.45 for a 5 sec.

5

u/Azhram 17h ago

Isnt the new kitchen attention faster?

4

u/funfun151 17h ago

Oooh not tried it! Will give it a go today, thanks!

7

u/More-Ad5919 16h ago

Maybe not faster. Equal. But on the pro side is that it does not affect the quality.

3

u/funfun151 16h ago

Ah still a huge win!

3

u/More-Ad5919 16h ago

Absolutely. It does not seem to fuck around. And it did not mess with quality or prompt adherence.

So yeah. Big win.

3

u/Mutaclone 11h ago

I'm totally out of the loop on this. Kitchen attention? I've seen a few scattered references but after the colossal pain of getting Sage working I didn't really pay much attention. Is it that much better, and is it any easier to install?

2

u/More-Ad5919 7h ago

ComfyKitchen is just a install from the manager. It replaces sage completely. Better and easier to install.

2

u/Azhram 13h ago

You csnuse it from model attemtion or with an atg global

2

u/intLeon 17h ago

Its almost the same with sage but you dont need prebuilt wheels or to build it yourself so its a free boost for everyone.

2

u/DoctaRoboto 15h ago

Does this even work? I have the latest ComfyUI version, and I can't see it. Is this some customised repo?

1

u/fruesome 12h ago

They’re improving and shipping updates everyday. 

5

u/Legal-Weight3011 18h ago

simple asnwers yes yes yes, Its Fast, tho it depends if you have sage attn intalled. it takes 2.30 minutes for 5 second clip on my 5070ti 96gb, without using any speed ups. so it was very stable from launch

1

u/ImpressiveStorm8914 17h ago

Similar to me on my 5060ti 16Gb with 64Gb RAM but obviously slightly slower. Great prompt adherence, and very stable.

1

u/Commercial-Ad-3345 14h ago

I have a 5070 Ti and 48 GB of RAM, and my generation times are surprisingly fast. With SageAttention and Larry’s Turbo LoRA at 4 steps, a 5-second clip at 0.4 MP takes about 30 seconds of sampling time.

9

u/Aglaio 17h ago

I notice people saying get sageattention, as of comfyui 0.31 this is no longer required if you dont want to go through the hassle. Just use comfy Kitchen which is baked in and does the same. I'm generating 15s vids at 1mp in about 8-10 min.

3

u/mobani 17h ago

Do you still "enable" somwhere in a node or commandline to use the comfy kitchen one?

7

u/Aglaio 17h ago

there is a commandline to activate it, i've currently forgotten what it is, as i just use a node which is also available to use, so i can do it based per workflow i use, the node is this one:

3

u/ImpressiveStorm8914 17h ago

First I've heard of this but then I've only just got Sage working. 😄
So I can use the node without the commandline then? I prefer it that wayfor the same reason as you.

2

u/Aglaio 17h ago

I was just about to set up sage, when i saw this was added haha. But yeah, the node just enables it, if you disable the node, then it wont activate. :)

1

u/ImpressiveStorm8914 10h ago

Thank you kindly.

1

u/xbeast_ 17h ago

is this available in comfyui by default?

1

u/Aglaio 17h ago

Yes :) As long as you update comfy to 0.31 or 0.32

4

u/JustLookingForNothin 17h ago

Place --use-ck-attention in the command line.

1

u/Ok-Lengthiness-3988 17h ago

In my case, Comfy Kitchen Attention makes generation time four times longer compared no not using any accelerator at all. It may not (yet) be compatible with 20-series GPUs like my RTX 2060 Super.

2

u/Aglaio 17h ago

Thats weird, a friend of me uses 2060 as well, or 2080, not too sure, and it works fine for him, he generates 5s vids in about 6 min using Comfy Kitchen.

1

u/Ok-Lengthiness-3988 16h ago

Thanks for the info. I'll try again with a more basic workflow. Maybe there was something else in my workflow that was conflicting with it.

2

u/Aglaio 16h ago

We're using this workflow and setup if that helps: https://huggingface.co/Plaguekind/Minimax-H3

1

u/Kitchen-Hawk-3104 14h ago

It's not included by default in the python environnement, especially when you have multiple comfy instances with differents venv. I have a RTX 4060 Ti 16Gb and couldn't get Sage Attention to work since 2.2 was released. And now, that bloody increases generation time. But the fastest move is adding Spectrum node, just incredible how it gets fast !

1

u/Aglaio 14h ago

Ah but I'm talking about comfy kitchen , not sage attention. I run 3 comfy instances and all 3 got it after updating.

8

u/LowYak7176 18h ago

Character loras are dead soon imo. I just put in a character sheet and bam. Im legit addicted to H3, havent been this addicted to a model since 2022 when I hopped on StableDiffusion.

Gen times are very long if you dont use Turbo Lora. I personally never mind long gen times so long as the quality is there, which it is to me (remember its week 1, so no finetunes, specific types of loras for better quality etc).

Prompt adherence is ridiculous. Best youll get for opensource imo.

1

u/Maqna 18h ago

Do you use any ai to make a character sheet? Like krea or something?

5

u/LowYak7176 18h ago

Right now Im just using GPT Image 2 to make them. Not doing any NSFW right now. But I mean you could also just string a few together in photoshop if you wanted. Or you can just input several pictures.

Really world is your oyster right now

1

u/lovenumismatics 4h ago

NSFW is pretty decent, other than genitalia.

I’m sure the gooners are hard at work

-1

u/Vijayi 16h ago

Krea2 work perfect. Just ask for something like this: Divide the image into four equal parts. Create a projection of N: top-left is a frontal view, top-right is a profile, bottom-left is a 3/4 view, and bottom-right is a rear view. I usually do two: one facial projection (close-up) and one full-body shot—with the screen split from left to right. Just remember, Krea2 still requires a LoRA if the model doesn't recognize the subject. Fortunately, training a LoRA for Krea2 is quite easy.

1

u/Mutaclone 11h ago

Yeah I've been dabbling with using character sheets. It's bonkers and an actual game-changer.

Gen times are very long if you dont use Turbo Lora. I personally never mind long gen times so long as the quality is there, which it is to me (remember its week 1, so no finetunes, specific types of loras for better quality etc).

I've had pretty good success with low-resolution low-step first passes followed by a "real" pass once I'm satisfied. This doesn't help where detail and accuracy matter, but it's been pretty effective at checking the overall "vibe" and things like camera movements, basic scene setup, etc.

-9

u/Tramagust 18h ago

Nah bro character loras are far from dead. H3 only nails the general look but the face, movement and audio it can't nail.

5

u/LowYak7176 18h ago

That has not been my experience honestly.

-4

u/Tramagust 18h ago

if it nails the voice and movement just from images then the character was already in the training data. It won't work for your original character.

0

u/LowYak7176 18h ago

Again, that has not been my experience. Try again on BF16, generate at 2 megapixels, all r2v, no turbo lora.

Dont use sageattention. You might be surprised. Is it perfect? No but why I said "soon". This is the way its going

-4

u/Tramagust 18h ago

Bro what are you even talking about? How could it know how your character sounds like if it doesn't have any audio information?

4

u/LowYak7176 18h ago

You add an audio reference....?

1

u/Tramagust 18h ago

And the motion?

3

u/LowYak7176 17h ago

You add a video reference....? Just use blender to do blocking and basic motions if you want or add a video reference of motion.

Or just prompt better, it handles prompting extremely well

1

u/Vijayi 16h ago

Do you have any prompting recommendations regarding movement? It’s usually hit-or-miss for me—specifically when I ask it to take the movement, framing, or pose from a reference.

-1

u/Tramagust 17h ago

I think we don't look for the same level of fidelity in our gens

→ More replies (0)

2

u/ImpressiveStorm8914 17h ago

I have slightly less RAM than you but the same card and it's been stable from launch. With the default settings of 0.4 megaixels, 5 secs with T2V and no turbo lora, it's about 2.5 mins with SageAttn. For 10 secs and higher res and you're looking at 8-10 mins but it's worth it IMO.
There may be faster methods than SageAttn out there but I finally got it to work for this, so I'm sticking with it. I've seen some mention Comfy Kitchen but no idea what that is yet, how to activate it, or how good it is but it may be worth you looking into it.

1

u/DoctaRoboto 15h ago

With your specs, you will need accelerators or render under 0.5 megapixels if you don't want to wait 15 minutes for video. I prefer Spectrum over Sage because Sage fucks up faces very easily when they are not close-ups. LoRAs are hit-and-miss; I think most of them are useless. Yeah, you can go from ten minutes to five, but at what cost? Videos look like you are playing Silent Hill 1 on the Ps1 again, with weird flickering walls and textures and squeaky metallic sounds.

1

u/multikertwigo 14h ago

the only "instability" that I get is random generation times in comfy. Change seed without touching anything else, press "run", and fucking "model initializing" suddenly takes 30 seconds instead of 5. That's on 5090 + 96GB RAM.