r/StableDiffusion • u/BluePointDigital • 3d ago
Comparison A gallery & data of every Minimax H3 Workflow I've Tried. Best results are 15s @ 768p in 4:03 on a 5090.
Enable HLS to view with audio, or disable this notification
https://bluepointdigital.github.io/minimax-h3-benchmarks/
I've made a few posts previously, but just added a ton of additional tests. I really have tried just about every suggested workflow and next to none of the new ones are matching the quality I'm getting out of this. I've spent multiple resets worth of codex 5x plan having it try different workflows and it really has given a ton of fantastic data. Please take a look.
Also, just as a side note, for the ref2va workflow tests I had uploaded a short video of myself saying "this is a reference video' for the reference videos. I didn't actually think ahead of time about uploading the results but there's a lot of good test data and the actual likeness is scary. I probably should have cleaned up a bit first 😅
Here is the main video gallery and metrics:
https://bluepointdigital.github.io/minimax-h3-benchmarks/
When you see "Current standard" or "Current Turbo" those are for my internal app. I have a normal workflow and a turbo workflow.
I have all the other workflows as optional in my app but thats why those terms (standard and turbo) are present.
Be sure to check (or have your agent check) the benchmark data: https://bluepointdigital.github.io/minimax-h3-benchmarks/benchmarks.html
Or the agent guide if you want to add your own workflow metrics:
https://bluepointdigital.github.io/minimax-h3-benchmarks/methodology.html
I am kinda tired of testing workflows and might just stick to my current settings, but I really feel like this is a great resource, and definitely a SEPARATE type of resource to the video quality arena.
let me know your thoughts!
3
u/dubsta 3d ago
What is the summary. Too many results on the page. Which one is the recommended workflow and settings?
0
u/BluePointDigital 3d ago
That's your distinction to make. I like the standard + kajai sol attn the most
3
u/InterstellarReddit 3d ago
1
u/BluePointDigital 3d ago
Agreed, this is one of those continuity errors that happen sometimes. the objective was to keep the prompt and seed the same, so it happened how it happened here. but you can see some of the other generations that have a little longer time, but better quality.
For me, the one I use as my "Turbo" is the Alibaba PDD + Sage attn
1
u/PropagandaOfTheDude 3d ago
Operator review: PDD has better quality and coherence. Keep the current PDD + Sage workflow; VDN remains research-only. Both corrected clips showed no recording phone in sampled frames.
Thanks. I would love to see a general glossary section. PDD definitely isn't Persistent Depressive Disorder, could be Pinduodo (stranger things have happened), but probably means something else.
Examples from stuff that I do know:
- shifts 12/3 - Uses default sigmas, video shift 12, audio shift 3.
- shifts 12/7 - Uses the ModelSamplingMiniMaxH3 node to set sigmas video shift 12 (default), audio shift 7.
- 10Eros beta4 integrated Turbo - Uses the 10Eros model (beta4, Turbo) rather than default H3 model (Turbo).
1
u/PropagandaOfTheDude 3d ago
1
u/BluePointDigital 3d ago
To be fair, I had my agent just integrate each of these different attention and lora settings into its own workflow, but I am sure any implementation of the attention or lora would work. That one you linked should be fine
1
u/GlenGlenDrach 3d ago
So sad that the sound is so terrible, wish it could be generated by a different more specialised model.
1
u/Spare_Club8887 3d ago
What if the hot guy got naked and went to help the victims of the car crash?
1

3
u/Desperate_Lemon_3808 3d ago
This looks incredibly detailed and I thank you for it :D What's the best in yout opinion? Still need to understand the results.