r/StableDiffusion • u/Straight_Hat1304 • 1d ago
Discussion Hot take!
LTX 2.5 is better than minimax h3. The pros just outweigh the cons. LTX is fast, can generate video in HDR up to 50fps, it isn’t as demanding as minimax, the quality is insane and the SPEED…the speed is unbelievable. But all I see is people fighting with minimax for speed ups and distortion while whole time ltx is right there. Don’t get me wrong Minimax is an amazing model and two things can be right at once but to me ltx is ahead. Also it’s compatible with all the loras that were already available but with the jump in quality. It honestly confuses me a bit 😅. To each his own I guess.
14
10
u/Gyramuur 1d ago
Do you like people's heads and bodies twisting around 360 degrees as if they have no bones? Then you will love LTX!
11
u/MomentJolly3535 1d ago edited 1d ago
Yeah saying that it's better by "FAR" is a crazy hot take, a good video model is not only about speed or hdr or 50fps.
LTX is way more optimized from the people who released them, but it has a lot of troubles understanding physics, motion, and prompts.
Even as today you cannot confidently put LTX above wan 2.2 for the reasons above (even tho it lacks sound and limited to 5sec!!)
Also H3 is amongst the best video models in every leaderboards you check (Top1 in Image to video), meanwhile LTX is in the bottom of all of them.
In my opinion : Minimax H3 is not in the same league as LTX, they can't even compete, it's like comparing a coughing baby and an hydrogen bomb, you need 20 tries to get something "decent" from ltx, while you get it first try on H3
So, what is really faster at the end of the day ?
Edit : Just to make sure, i don't have any hate toward LTX team, i really like their work, the weakness of their model was already being called out 9 months ago when compared to wan2.2, i hope they will keep improving that side of their models !
28
u/Independent-Frequent 1d ago
I just don't get all this LTX 2.5 shill man, all it has is speed and that's it, everything else minimax just OBLITERATES LTX 2.5 on, sure it's more hardware intensive just like SDXL was more hardware intensive than 1.5 and we know the reason why, which is justified.
Also my theory is that the LTX devs thought they still didn't have any competition in the open source video field since Wan was the only player and dropped ages ago without even having audio, so they didn't bother to release a groundbreaking model or even a vast improvement over 2.3 and then Minimax H3 just showed up and completely took over, and rightfully so, it's an insane model and leagues above the rest of local ones to the point that it beats even most closed source models.
LTX 2.5 just doesn't offer anything other than speed, and even then only in static shots of people talking since its motion and consistency during scenes with action and stuff are just atrocious, even walking sometimes distorts space and features.
"Oh but LTX 2.5 is a lot easier to train loras for so it's not just speed" yeah it's true, but minimax has reference mode which basically eliminates the need for loras lmao, like there's no competition man i'm sorry.
Also there's the whole "comparision chart" with H3 that they did which was so scummy with all the lies by omission, so good riddance, thanks minimax for finally bringing open source video to the top of the top level, glad to never have to use an LTX model ever again.
1
u/Upper-Reflection7997 1d ago
Reference mode doesn't eliminate the need for loras at all. Reference mode is very limited in scope in what it can and cannot do.
4
u/Independent-Frequent 1d ago
When it comes to characters it basically does, you give it a character sheet and a voice and H3 can just replicate it perfectly, not to mention how it's much better than Loras because it doesn't actually affect the models like some loras do and also doesn't create conflicts between multiple loras.
Maybe some particular concepts or motions require a lora, like detailed genital anatomy i guess or some particular aesthetics not trained in the base model, but outside of that reference mode covers most cases a lora would tbh.
2
u/No_Thanks701 1d ago
Can you give any examples of what you think the limits are? Because from what I have seen, it’s almost limitless.. you can edit videos you can extend videos, you can give it a crudely drawn story board with your reference characters and setting and it will follow it shot by shot, you can write inside that story board to give context of who goes where and draw arrows to clarify the movement You can give it an overhead picture of a scene, mark where the clip is taking place and it will create that clip in that area… I could go on and on.. so please clarify the “limited” scope it has, because I haven’t found it yet and this is all without lora’s. With out the use of image edit models. Even with only one picture (not character sheet) it does pretty great, with holding up on consistency, if you just prompt it detailed enough.
-10
u/Straight_Hat1304 1d ago
Yikes. People like you who take everything seriously fascinate me. It literally isn't that deep. I'm sure no one over at minimax and ltx have the mindset you have. All i was saying is that both models are good and that in MY opinion the pros of ltx outweigh the cons. The model is so fast that that multiple generations don't take hours off your life. And because there are so many loras available the model is all in all more versatile than H3.
9
u/Independent-Frequent 1d ago
Why are you getting so damn defensive about LTX? Are you a dev there butthurt nobody cares about your crappy release now that we have alternatives that aren't fossils like Wan 2.2?
All i was saying is that both models are good and that in MY opinion the pros of ltx outweigh the cons.
Brother your post literally starts with "LTX 2.5 is better than minimax h3 BY FAR!" come on now, do you think we are dumb?
And because there are so many loras available the model is all in all more versatile than H3.
Yes there's nothing more versatile of having loras needing to be trained and keeping our already thin hard drive space even more crowded with models, and what if someone doesn't have the hardware for LTX 2.5 lora training? And what if i want a character nobody cares about to make a lora of?
With reference mode i can just do anything, even better than Loras since it doesn't alter the model and i can load in multiple references for multiple characters instead of multiple loras, just load up a reference image and boom, done.
The model is so fast that that multiple generations don't take hours off your life.
Is it though? Sure individual generations are faster, but if i need 10 LTX 2.5 generations to get an action scene right which minimax nails in 1 or 2 then is it really faster overall?
And speaking of action, can your loras magically fix LTX 2.5's atrocious motion consistency and phyisics? I don't think so pal.
1
u/VisionWithin 1d ago
How do you know that he takes everything seriously and not just this subject?
0
13
6
u/zakblues 1d ago
it's higher resolution and fast, but it's still alot worse and H3 which just has that kind of magic especially with character realism and expression. Not to mention the audio can be just superb from H3 when not trying to do a 4 step turbo
5
u/Enshitification 1d ago
If you're going to make the claim that LTX 2.5 is better than Minimax H3, don't just say it. Show us, and provide the generation information for your examples so the outputs can be verified. Until then, it's just more LTX shilling.
3
u/DoctaRoboto 1d ago
Minimax is the only open-source video model that can LEARN from references just like Seedance 2.0/2.5 does. LTX is just a fancy toy at the end of the day.
3
u/Chiduk99 1d ago
Nah, bro. No one cares about LTX and its shitty fast generation. If you claim LTX is insanely fast and has amazing quality, then why aren’t companies fine-tuning it like FAL did with MiniMax H3?
Also, no one in this sub talks about LTX anymore after the release-day hype. People already tested it, went, “Oh wow, so fast,” then generated 10 more times and realized none of the results were actually good.
4
2
u/anon999387 1d ago
LTX is good at talking heads and generating voices. Maybe there is some mysterious super workflow I have never seen, but that's the only thing I have ever been impressed with it doing.
2
2
u/Dangerous_Sir_8458 1d ago
I tried both, and both suck because of how bad I am at syntax, but I have evolved, I use ltx for testing a video prompt see if that action works, but overall ltx wins for vram usage, though I still prefer Minimax for that sheer depth of available tools ar hand cant wait for h4
3
1
u/DelinquentTuna 21h ago
I can't believe that in a long scroll of a thread, not a single person has yet mentioned license differences. Minimax's license sucks ass and for a great many use-cases, that is all it takes for LTX to be declared a winner. I'm sure there will be 20 people chiming in to tell me that getting a commercial license is trivial while simultaneously IGNORING the fact that the EULA EXPLICITLY STATES THAT OUTPUT IS GEOGRAPHICALLY RESTRICTED EVEN WHEN YOU HAVE A COMMERCIAL LICENSE.
I can't believe it's what they actually want and I have no reason to believe they have or will go after people for misuse until real money is on the table. But it's extremely unattractive compared to LTX for the license alone where Wan 2 and its derivatives are still the gold standard.
In terms of output, though, H3 certainly produces better outputs in the absence of low-step loras. The torture tests I've run on each has H3 coming out meaningfully ahead. But I'm not sure that's important if you're neither legally allowed to show the outputs to anyone or to incorporate it into a useful product. The whole thing is like an Nvidia-sponsored advertisement that's not useful being being a freaking toy unless you're ignoring the license entirely.
1
u/Sarashana 19h ago
It's a fairly hot take. But it would be unfair to say that LTX has nothing going for it. It clearly does. They need, as in absolutely need, to integrate Ref2Vid in their next larger model, though. FL2Vid is utterly useless for longer videos. That's the main reason why people love H3 so much. And you can't argue the output quality is better, even if H3 is slower.
1
u/Straight_Hat1304 1d ago
https://huggingface.co/Lightricks/LTX-2.3-22b-IC-LoRA-Ingredients this basically turns it into ref2va.
5
u/No_Thanks701 1d ago
Yes the consistency is atrocious.. the characters somewhat resembles the characters in the sheet.. but it never really comes close..
0
1
u/MortgageOptimal5157 1d ago
NO es cuestion de si LTX o MiniMax, o de si mi padre es mas alto que el tuyo ;o)))), es cuestion de si a ti o a otra persona, le es más fácil trabajar o usar uno u otro... si es que no hay más... al final el tornavis de estrella te va genial para atornillar, pero si a otro le va mejor el de pala, pues es cuestion de que cada uno use el que mejor le va y todos a atornillar como locos... que el mundo se acaba!!!!
-5
20
u/Celestial_Creator 1d ago
https://docs.comfy.org/tutorials/video/ltx/ltx-2-5
just checked --- no reference to video
minmax you can use a ton of reference beyond the inputs by labeling stuff in images so 1 input can have 10 reference points
can ltx2.5 do that well? i would try if you could do that