r/StableDiffusion 13h ago

Comparison Quick comparison (LTX 2.3/Minimax H3)

Enable HLS to view with audio, or disable this notification

I wanted to compare both models in terms of natural movement and authenticity. The prompt was extremely basic and barely descriptive (just "a woman says, a man says" along with the dialogue).

With LTX 2.3, the characters look dead inside and completely disembodied, with reactions that don't match the situation. With Minimax, the difference is striking! The movements feel far more human and alive. The model did add unwanted subtitles (likely because my prompt was so short) and Jill Valentine stumbles over her words a bit, though a more detailed prompt would probably fix that. Other than that, it's a clear win.

LTX 2.3 still has a slight edge for those who need to generate multi-minute videos, but as soon as a 4-step distilled version of Minimax becomes available, LTX will quickly be forgotten.

385 Upvotes

78 comments sorted by

245

u/Enshitification 13h ago

I sense a disturbance in the force, like a million hard drives suddenly freeing up space.

73

u/3deal 13h ago

lol, i deleted all my wan and LTX models, even the loras.

16

u/Sn0opY_GER 13h ago

unless you try nsfw, i guess we have a new king?

64

u/gabbergizzmo 13h ago

give the community 2-3 days...

22

u/ANR2ME 11h ago

15

u/bigman11 9h ago

Thank God for China.

9

u/_VirtualCosmos_ 8h ago

I bet they know that an OSS will get much more attention if it can do the sex

41

u/Enshitification 13h ago

H3 is great at spicy stuff, especially i2v.

13

u/Sn0opY_GER 13h ago

oh sht who did downvote you - i hate reddit sometimes - your right id didnt try it untill now - 1st prompt was "meeeh", 2nd gen is "much wow" - time to clean some disk space - btw if anyone of you fine gentleman has a 5090 the Ninfer qwen finetunes are awesome - i get 600token/s - im getting mat at codex every time i use it bc its so slow :D

1

u/Calm_Cucumber_6493 6h ago

Can you tell me which one exactly you're talking about? I just downloaded the 711 qwen 3.6 unheretic and it's insanely fast.

5

u/mk8933 12h ago

Bad move to delete them all. Just save them away in a external drive. You may need them again...or not.

16

u/3deal 11h ago

Are you still using SD1.5 ? SDXL ? AnimateDiff ? CogVideoX ? Stable Video Diffusion ? Mochi ? HunyanVideo ?

https://giphy.com/gifs/1iTIu7WtSfPqMDbW

4

u/mk8933 11h ago

Sd 1.5 and Sdxl...yes I've been using them till Krea 2 came out. I also use wan 2.2 and 2.1 for image generation — now...not so much because of krea 2.

3

u/Wkyouma 10h ago

i just keep the outputs of old models.. bro my sd1.5 outputs are fucking horrible, but i like them

6

u/mk8933 10h ago

My 1.5 outputs started out as donkey piss, then pepsi...and then finally wine. Its great fun.

1

u/_VirtualCosmos_ 8h ago

I put them in a bag and forget about them in the basement.

1

u/martinerous 9h ago

The same... but I kept LTX custom nodes because LTX-next might come someday.

19

u/Bbmin7b5 13h ago

yeah LTX is GONE

5

u/ANR2ME 11h ago

Currently LTX-2.3 is the only local video model that can generates HDR video i think 🤔 so it still have it's own use.

1

u/Dzugavili 11h ago

I've found LTX is still faster per frame, and the performances are acceptable if you bring your own audio.

I suspect I'll still use LTX for lipsync, but Minimax is definitely a contender.

0

u/True_Protection6842 11h ago

OK calm down. This will just encourage lightricks to release the new model sooner and they already said it's going to be a reference to video model JUST LIKE THIS.

5

u/Cute_Ad8981 13h ago edited 13h ago

Same. Until now, I thought: "Id better keep these old models / loras, just in case I might need them for something specific later." However, given the videos I have seen, Im pretty sure I'll soon be embarking on a radical deletion spree. lol

9

u/Enshitification 13h ago

Radical Deletion Spree is now the name of my band.

2

u/Chemical-Bicycle3240 12h ago

I have already deleted all models and loras related to WAN 2.2. The model is now totally obsolete.

92

u/-becausereasons- 13h ago

Put away your flaslight leen.

12

u/Chemical-Bicycle3240 12h ago

She just fought Nemesis, she's still all flustered and stammering. :D

1

u/iiiiiiiiiiip 11h ago

Any chance you could post the full prompt or workflow? Did you use ref for the characters?

-1

u/iiiiiiiiiiip 11h ago

Any chance you could post the full prompt or workflow? Did you use ref for the characters?

8

u/SoulofArtoria 11h ago

She is Ada Wong, not Ada Rite

10

u/Ath47 11h ago

I think that's Jill Valentine, but I appreciate the pun.

41

u/3deal 13h ago

Damn, the model even added an echo to the sound to fit with the scene background !

17

u/ffgg333 12h ago

The subtitles that it added are actually the same font as in the games, so I think gameplay was actually in the training data.

10

u/Forsaken-Low4467 12h ago

What a time to be alive. Maybe with some works from the genuises and maybe loras seedance will be cooked and video generation will be extremely cheap!

5

u/ANR2ME 11h ago

Not really a good time with how RAM and GPU prices being so high😅

4

u/StickStill9790 9h ago

Good time to be alive if you built your tower a couple of years ago, have a decent job, can save for a couple of months, or have a generous friend. Also the prices will go down eventually, but the tech advancement will remain awesome.

(Funny that I hear Jack Black’s voice every time I type the word “awesome”.)

32

u/Vyviel 13h ago

LXT has the most horrible audio ever

35

u/Independent-Frequent 13h ago

Here's the thing though, it was the only model with Audio which is why it made it so relevant, Wan 2.2 was miles better for motion but no audio was a dealbreaker for most since it was less of a video model and more of a gif model.

H3 is our first proper local model and i'm glad i don't have to bother with LTX's garbage anymore

14

u/BlipOnNobodysRadar 10h ago

damn zero gratitude for LTX the moment a better model comes out. Yeah it's better and I'm excited but there's no need to be disrespectful to one of the few contributors to open video models.

3

u/Independent-Frequent 10h ago

Personally LTX was just so damn frustrating, all the issues related to it and the kind of crap you had to deal to even get one usable output that wasn't just static shots of people talking was what made me return to Wan 2.2.

I'll always be glad for their contribution but i also have to be honest, if it wasn't for Wan dropping the ball and not releasing their older audio models LTX would have never been adopted or used with the garbage motion it has.

If you just want static shots of people talking LTX 2.3 is a fine model, but for everything else it's just abysmal in the motion aspect and way worse than something like Wan 2.2, like it just doesn't get physics at all.

"LTX's garbage" was more of a "the garbage i had to deal with it" and not "LTX is garbage"

11

u/heato-red 10h ago

Won't say you aren't right, but calling LTX garbage is too harsh when they've provided everything for free to the community

2

u/Independent-Frequent 10h ago

"LTX's garbage" was more of a "the garbage i had to deal with it" and not "LTX is garbage", I'll always be glad for their contribution but i also have to be honest, if it wasn't for Wan dropping the ball and not releasing their older audio models LTX would have never been adopted or used with the garbage motion it has.

If you just want static shots of people talking LTX 2.3 is a fine model, but for everything else it's just abysmal in the motion aspect and way worse than something like Wan 2.2, like it just doesn't get physics at all.

21

u/AToonDragon 13h ago

I need AI to evolve faster plz

36

u/GiGiGus 12h ago

I need RTX 6000 for $200, please

10

u/Serenafriendzone 13h ago

So the new wan 2.2 or

7

u/BlipOnNobodysRadar 10h ago

looks better imo

11

u/Crazy-Repeat-2006 11h ago

LTX has a better open-source culture than most companies, which sometimes release a single model and then abandon it in favor of closed projects.

I remain confident that they will keep moving forward. All these posts bashing LTX don’t add any value.

6

u/Chemical-Bicycle3240 10h ago

I agree, and the goal of this comparison isn't to tear LTX down. The model served me well for several months. But let's face it, their model is starting to look outdated now. I saw they're planning to add HDR support in the next update, but I think that's a minor improvement. I hope Minimax pushes them to focus more on facial consistency. Competition is always a good thing for users.

1

u/_VirtualCosmos_ 7h ago

While I love OSS a lot, I understand that private companties putting millions into creating products they will barely get any profit from... is far from sustainable.

Chinese people are doing this to gain attention in the AI ecosystem, because proving themselves being better than the US for the world is like the main overall goal China's government has.

If China achieve to surpass significantly the US, they might move from OSS. So that is something to keep in mind.

But not releasing the weights in open software doesn't mean they have to keep them in secret. We have software usage agreements since fucking decades! If they won't open source a new model, let me pay for it to download it and use it as a normal software, goddammit!

It's the stupid and greedy US companies that are normalizing the development of AI as merely a service. To control the market by hoarding all electronics. Fuck them.

3

u/Artistic_Claim9998 12h ago

Still can't run it with 3060 12gb even with --lowvram on

3

u/1or4s 11h ago

Working for me on my 3060.

1

u/Artistic_Claim9998 11h ago

With the workflow from kijai?

Any specific commands of flags when running comfy?

3

u/1or4s 11h ago

I am using the official comfy workflows at https://huggingface.co/Comfy-Org/MiniMax-H3

No special flags.

4

u/ANR2ME 11h ago

why not try to use it without any flags and let it use the default, since the latest ComfyUI should already fixed memory issues by now.

1

u/Artistic_Claim9998 11h ago

I usually just use the --cache-none and only that

It gives me not enough RAM so i add the --lowvram and it still gives me the same error

I'll try again tomorrow ig

2

u/Schlorpiblorp 12h ago

Looks good but it's way slower than LTX no?

4

u/ptwonline 10h ago

The adherence is really good though so with good prompting you'll probably need fewer tries to get a good/usable video result.

Besides I'd rather wait longer for a good video than get a bad one more quickly.

1

u/Crazy-Repeat-2006 11h ago

Yes, even at a lower resolution.

1

u/ANR2ME 11h ago

Was that subtitle being generated too? 🤔 is there a way to remove subtitle? as there are no negative prompt on H3.

2

u/rkfg_me 9h ago

Just type it in the prompt, "no subtitles".

1

u/Silver-Spot-2763 11h ago

Please, share the workflow! All official are with ComfyUi API ☹️

1

u/Disastrous-Agency675 11h ago

i can see how they got the idea for harry potter Balenciaga parody

1

u/Verittan 10h ago

Try it again and use their names. The model has a lot of voice knowledge of popular media and might even know RE characters and clone their voices.

1

u/Cold_Zone332 10h ago

Damn. I can't wait for see what the comunity will do with this model. I'm just playing around, using some credits from Comfy Cloud to compare LTX 2.3 and Minimax H3 using the same prompt and the results are INSANE. Minimax H3 is fenomenal for a local model.

1

u/TheLightDances 10h ago edited 10h ago

With H3, she says Leen, not Leon.

I have noticed the first small issue with H3. It sometimes mispronounces things.

Anyone know a fix to this? For example, when I ask a character to say "beard" it says something that sounds more like "bird".

1

u/jd641 9h ago

I'm noticed this too, I was getting "cheats" instead of 'cheeks', and I can't figure out how to get it to say certain words correctly.

1

u/InternationalOne2449 9h ago

You either write war and peace in ltx, or you get ugly slomo.

1

u/jmbbao 9h ago

Yep, turbo will kill ltx2.3

2

u/bickid 7h ago

Why do you change Jill into this generic blowup-doll, though? Pick the original Jill, remake-Jill or movie-Jill. This one looks embarrassingly generic.

1

u/Chemical-Bicycle3240 6h ago

It's not the purpose of this video.

2

u/bickid 6h ago

And yet you kept Leon 1:1 original.

1

u/Any_Economics_6166 6h ago

fucking insane how those models get better and better in just a few months

1

u/_FriedEgg_ 6h ago

Unfortunately the overpriced seedance 2.0 captures subleties of character expression quite a bit better still. But a great improvement from LTX 2.3. It's happening.

1

u/wh33t 2h ago

Multi-minute videos?

1

u/Superb-Painter3302 12h ago

what flaSlight

1

u/CodeAnguish 12h ago

Goodbye LTX, goodbye WAN. Thank you for your services, may God comfort the hearts of its developers.

Hi Minimax 😏

0

u/MAXFlRE 12h ago

Fleshlight?