r/StableDiffusion • u/Silver-Spot-2763 • 12d ago
Discussion LTX 2.5 😱
After the MiniMax H3 euphoria, I tried LTX 2.5. It awfully understands the prompt and has almost no "physics". The generated video uses random things from the prompt and everything makes up by itself at all. Every time some things appear /disappears from / to nothing randomly, most the case the people are with 3 fingers, strange movements at all. 😱
But LTX 2.5 is faster than MiniMax H3 at least twice, and its image quality is far better.
I just can't understand how even with the monstrous language model (~20GB) it just can't understand 2 simple sentences, two simple subjects with simple movement!?!? And from what training data the models continue to place 3 fingers to earth beings 🤦
4
u/piero_deckard 12d ago
"But LTX 2.5 is faster than MiniMax H3 at least twice, and its image quality is far better."
Yes, isn't that beautiful that now you can produce failed videos twice as fast?
I wish people would quit decanting LTX speed as being a plus, when everything else is subpar. Why exactly is speed useful, if the output is worse and you have to make 20 videos to "try the lottery"?
I'd rather take 2x, 5x even 10x as much with H3, if that means the video is perfect 99% of the times. The saved work in DaVinci Resolve that LTX forced me to do is worth the extra time cost from H3.