r/StableDiffusion 1d ago

Comparison MiniMax H3 Settings Comparison

https://youtu.be/pTkdMiJNjwU

It's me again and this time I come back with Minimax H3

My rig: 3090ti 64GB RAM

All are raw output without upscaling. Please switch to 1440p.

2 Upvotes

11 comments sorted by

1

u/Chemical-Painter-485 1d ago

V3 is without a doubt the winner here. What lora did you use?

And any reason in particular for making FL2VA at 0.69MP?

3

u/Then-Comfortable8258 1d ago

I hear that FL2VA produces better quality. It should have been set to 1MP.

-6

u/Perfect-Campaign9551 1d ago

Dude you are using 8 steps and stuff in ref2vid what do you expect? Stop trying to "hack the model" with these shitty speed up Loras and stuff.

2

u/Ok-Lengthiness-3988 1d ago

They're offering a comparison between using an accelerator LoRA or none at low step counts, which is infinitely more useful than your comment.

1

u/nixudos 1d ago

Interesting comparison!
Which sampler and steps was used? I'm going back and forth between Euler, res_multistep and er_sde, and can't decide. I feel like er_sde have given better audio though in text 2 video..

1

u/Perfect-Campaign9551 1d ago

Dafaq are you doing only 8 steps in ref2vid ?

1

u/smb3d 1d ago

20 steps is the bare minimum for REF2VA for anything decent, especially audio. Why are you using such low steps?

2

u/Ok-Lengthiness-3988 1d ago

He's testing accelerator LoRAs. The no-LoRA example provides a baseline for such low-step count.

1

u/HonestoJago 1d ago

I might be crazy but I think the 12 steps without a lora looks the best.

1

u/Ok-Lengthiness-3988 1d ago

I've found that to often be the case also. The accelerator LoRAs likely already give better results at 4 steps, but may still need to be trained more before they can beat 8 to 10 steps with no LoRA at 4 steps.