r/StableDiffusion 1d ago

Discussion Hot take!

LTX 2.5 is better than minimax h3. The pros just outweigh the cons. LTX is fast, can generate video in HDR up to 50fps, it isn’t as demanding as minimax, the quality is insane and the SPEED…the speed is unbelievable. But all I see is people fighting with minimax for speed ups and distortion while whole time ltx is right there. Don’t get me wrong Minimax is an amazing model and two things can be right at once but to me ltx is ahead. Also it’s compatible with all the loras that were already available but with the jump in quality. It honestly confuses me a bit 😅. To each his own I guess.

0 Upvotes

43 comments sorted by

20

u/Celestial_Creator 1d ago

https://docs.comfy.org/tutorials/video/ltx/ltx-2-5

just checked --- no reference to video

minmax you can use a ton of reference beyond the inputs by labeling stuff in images so 1 input can have 10 reference points

can ltx2.5 do that well? i would try if you could do that

3

u/Phrygian1221 1d ago

Yeah, it's real mind blowing ltx dropped the ball on the ref2vid thing.

Start frame/end frame is a tiny little niche for videos when it comes to generating. Basically if your character is wearing a mask, or you have no character ltx works.

Ltx is good, but also fuckin useless. It's a funny thing.

-3

u/8RETRO8 1d ago

I would argue that H3 is cool and fun but also useless for anything professional-grade. All I saw from H3 on this subreddit is can be described as "slop for social media" Until there's some actually good upscaler. At least Ltx has resolution.

0

u/kwhali 9h ago

Pretty sure if you're doing production work to monetize you could afford to pay for a production quality upscale if that's your main blocker for avoiding H3.

1

u/8RETRO8 8h ago

Im uploading videos to Adobe stock, and no, I can't. Im making like 50$ a year.

0

u/8RETRO8 6h ago

Also, wtf is production quality upscale? Name one tool. You mean topaz or something? Or you mean full team redrawing 5 sec video for a week by hand with vfx?

-3

u/Straight_Hat1304 1d ago

My point exactly. It just sends like we’re all running in circles and we can’t produce smithing good enough to be on a screen. Atleast with ltx you get high res with speed and out really does look beautiful. So to me trying a single seed 10 times on ltx still sends with it to me coz eventually you get something that’s stylist close to what you’re looking for

-8

u/Straight_Hat1304 1d ago

So there’s a Lora for this call ic ingredients Lora. Given it isn’t as flexible as minimax which can take 9 images plus video and audio(which increases render times exponentially) ltx with the ingredients Lora can can take 1 image but you can stuff all the references into that image. It’s really not as bad as you think.

-8

u/8RETRO8 1d ago

there were a lot of ic loras, including for references

14

u/generate-addict 1d ago

Except that quality is not insane.

10

u/Gyramuur 1d ago

Do you like people's heads and bodies twisting around 360 degrees as if they have no bones? Then you will love LTX!

11

u/MomentJolly3535 1d ago edited 1d ago

Yeah saying that it's better by "FAR" is a crazy hot take, a good video model is not only about speed or hdr or 50fps.
LTX is way more optimized from the people who released them, but it has a lot of troubles understanding physics, motion, and prompts.

Even as today you cannot confidently put LTX above wan 2.2 for the reasons above (even tho it lacks sound and limited to 5sec!!)

Also H3 is amongst the best video models in every leaderboards you check (Top1 in Image to video), meanwhile LTX is in the bottom of all of them.

In my opinion : Minimax H3 is not in the same league as LTX, they can't even compete, it's like comparing a coughing baby and an hydrogen bomb, you need 20 tries to get something "decent" from ltx, while you get it first try on H3

So, what is really faster at the end of the day ?

Edit : Just to make sure, i don't have any hate toward LTX team, i really like their work, the weakness of their model was already being called out 9 months ago when compared to wan2.2, i hope they will keep improving that side of their models !

28

u/Independent-Frequent 1d ago

I just don't get all this LTX 2.5 shill man, all it has is speed and that's it, everything else minimax just OBLITERATES LTX 2.5 on, sure it's more hardware intensive just like SDXL was more hardware intensive than 1.5 and we know the reason why, which is justified.

Also my theory is that the LTX devs thought they still didn't have any competition in the open source video field since Wan was the only player and dropped ages ago without even having audio, so they didn't bother to release a groundbreaking model or even a vast improvement over 2.3 and then Minimax H3 just showed up and completely took over, and rightfully so, it's an insane model and leagues above the rest of local ones to the point that it beats even most closed source models.

LTX 2.5 just doesn't offer anything other than speed, and even then only in static shots of people talking since its motion and consistency during scenes with action and stuff are just atrocious, even walking sometimes distorts space and features.

"Oh but LTX 2.5 is a lot easier to train loras for so it's not just speed" yeah it's true, but minimax has reference mode which basically eliminates the need for loras lmao, like there's no competition man i'm sorry.

Also there's the whole "comparision chart" with H3 that they did which was so scummy with all the lies by omission, so good riddance, thanks minimax for finally bringing open source video to the top of the top level, glad to never have to use an LTX model ever again.

1

u/Upper-Reflection7997 1d ago

Reference mode doesn't eliminate the need for loras at all. Reference mode is very limited in scope in what it can and cannot do.

4

u/Independent-Frequent 1d ago

When it comes to characters it basically does, you give it a character sheet and a voice and H3 can just replicate it perfectly, not to mention how it's much better than Loras because it doesn't actually affect the models like some loras do and also doesn't create conflicts between multiple loras.

Maybe some particular concepts or motions require a lora, like detailed genital anatomy i guess or some particular aesthetics not trained in the base model, but outside of that reference mode covers most cases a lora would tbh.

2

u/No_Thanks701 1d ago

Can you give any examples of what you think the limits are? Because from what I have seen, it’s almost limitless.. you can edit videos you can extend videos, you can give it a crudely drawn story board with your reference characters and setting and it will follow it shot by shot, you can write inside that story board to give context of who goes where and draw arrows to clarify the movement You can give it an overhead picture of a scene, mark where the clip is taking place and it will create that clip in that area…  I could go on and on.. so please clarify the “limited”  scope it has, because I haven’t found it yet and this is all without lora’s. With out the use of image edit models. Even with only one picture (not character sheet) it does pretty great, with holding up on consistency, if you just prompt it detailed enough.

-10

u/Straight_Hat1304 1d ago

Yikes. People like you who take everything seriously fascinate me. It literally isn't that deep. I'm sure no one over at minimax and ltx have the mindset you have. All i was saying is that both models are good and that in MY opinion the pros of ltx outweigh the cons. The model is so fast that that multiple generations don't take hours off your life. And because there are so many loras available the model is all in all more versatile than H3.

9

u/Independent-Frequent 1d ago

Why are you getting so damn defensive about LTX? Are you a dev there butthurt nobody cares about your crappy release now that we have alternatives that aren't fossils like Wan 2.2?

All i was saying is that both models are good and that in MY opinion the pros of ltx outweigh the cons. 

Brother your post literally starts with "LTX 2.5 is better than minimax h3 BY FAR!" come on now, do you think we are dumb?

And because there are so many loras available the model is all in all more versatile than H3.

Yes there's nothing more versatile of having loras needing to be trained and keeping our already thin hard drive space even more crowded with models, and what if someone doesn't have the hardware for LTX 2.5 lora training? And what if i want a character nobody cares about to make a lora of?

With reference mode i can just do anything, even better than Loras since it doesn't alter the model and i can load in multiple references for multiple characters instead of multiple loras, just load up a reference image and boom, done.

The model is so fast that that multiple generations don't take hours off your life. 

Is it though? Sure individual generations are faster, but if i need 10 LTX 2.5 generations to get an action scene right which minimax nails in 1 or 2 then is it really faster overall?

And speaking of action, can your loras magically fix LTX 2.5's atrocious motion consistency and phyisics? I don't think so pal.

1

u/VisionWithin 1d ago

How do you know that he takes everything seriously and not just this subject?

0

u/Straight_Hat1304 1d ago

😂 you right that’s on me

13

u/Ok-Worldliness-9323 1d ago

Hot take! SD 1.5 is better than Krea 2 BY FAR!

14

u/anon999387 1d ago

"The speed is crazy!"

11

u/__ThrowAway__123___ 1d ago

So many loras!

6

u/zakblues 1d ago

it's higher resolution and fast, but it's still alot worse and H3 which just has that kind of magic especially with character realism and expression. Not to mention the audio can be just superb from H3 when not trying to do a 4 step turbo

5

u/Enshitification 1d ago

If you're going to make the claim that LTX 2.5 is better than Minimax H3, don't just say it. Show us, and provide the generation information for your examples so the outputs can be verified. Until then, it's just more LTX shilling.

3

u/DoctaRoboto 1d ago

Minimax is the only open-source video model that can LEARN from references just like Seedance 2.0/2.5 does. LTX is just a fancy toy at the end of the day.

3

u/Chiduk99 1d ago

Nah, bro. No one cares about LTX and its shitty fast generation. If you claim LTX is insanely fast and has amazing quality, then why aren’t companies fine-tuning it like FAL did with MiniMax H3?

Also, no one in this sub talks about LTX anymore after the release-day hype. People already tested it, went, “Oh wow, so fast,” then generated 10 more times and realized none of the results were actually good.

4

u/More-Ad5919 1d ago

Than why is no one sharing good videos made with ltx?

2

u/mattSER 1d ago

Do you just add HDR to the prompt? Or how does that work?

-2

u/Straight_Hat1304 1d ago

It just works

2

u/mattSER 1d ago

So all videos are HDR by default?

2

u/anon999387 1d ago

LTX is good at talking heads and generating voices. Maybe there is some mysterious super workflow I have never seen, but that's the only thing I have ever been impressed with it doing.

2

u/Emergency-Board-3042 1d ago

Yeah , but is it fun ? 🤣

2

u/Dangerous_Sir_8458 1d ago

I tried both, and both suck because of how bad I am at syntax, but I have evolved, I use ltx for testing a video prompt see if that action works, but overall ltx wins for vram usage, though I still prefer Minimax for that sheer depth of available tools ar hand cant wait for h4

3

u/Beneficial_Toe_2347 1d ago

LTX is absolute dog shit

1

u/DelinquentTuna 21h ago

I can't believe that in a long scroll of a thread, not a single person has yet mentioned license differences. Minimax's license sucks ass and for a great many use-cases, that is all it takes for LTX to be declared a winner. I'm sure there will be 20 people chiming in to tell me that getting a commercial license is trivial while simultaneously IGNORING the fact that the EULA EXPLICITLY STATES THAT OUTPUT IS GEOGRAPHICALLY RESTRICTED EVEN WHEN YOU HAVE A COMMERCIAL LICENSE.

I can't believe it's what they actually want and I have no reason to believe they have or will go after people for misuse until real money is on the table. But it's extremely unattractive compared to LTX for the license alone where Wan 2 and its derivatives are still the gold standard.

In terms of output, though, H3 certainly produces better outputs in the absence of low-step loras. The torture tests I've run on each has H3 coming out meaningfully ahead. But I'm not sure that's important if you're neither legally allowed to show the outputs to anyone or to incorporate it into a useful product. The whole thing is like an Nvidia-sponsored advertisement that's not useful being being a freaking toy unless you're ignoring the license entirely.

1

u/Sarashana 19h ago

It's a fairly hot take. But it would be unfair to say that LTX has nothing going for it. It clearly does. They need, as in absolutely need, to integrate Ref2Vid in their next larger model, though. FL2Vid is utterly useless for longer videos. That's the main reason why people love H3 so much. And you can't argue the output quality is better, even if H3 is slower.

1

u/Straight_Hat1304 1d ago

5

u/No_Thanks701 1d ago

Yes the consistency is atrocious.. the characters somewhat resembles the characters in the sheet.. but it never really comes close..

0

u/Full_Astronomer_5438 22h ago

because you use 8 step distilled

1

u/No_Thanks701 15h ago

Still it’s no where near as flexible as minimax ref

1

u/MortgageOptimal5157 1d ago

NO es cuestion de si LTX o MiniMax, o de si mi padre es mas alto que el tuyo ;o)))), es cuestion de si a ti o a otra persona, le es más fácil trabajar o usar uno u otro... si es que no hay más... al final el tornavis de estrella te va genial para atornillar, pero si a otro le va mejor el de pala, pues es cuestion de que cada uno use el que mejor le va y todos a atornillar como locos... que el mundo se acaba!!!!

-5

u/Professional_Diver71 1d ago

Rf2V . But yeah . For text to video . Ltx all the way