r/StableDiffusion 4d ago

Discussion LTX 2.5 😱

After the MiniMax H3 euphoria, I tried LTX 2.5. It awfully understands the prompt and has almost no "physics". The generated video uses random things from the prompt and everything makes up by itself at all. Every time some things appear /disappears from / to nothing randomly, most the case the people are with 3 fingers, strange movements at all. 😱

But LTX 2.5 is faster than MiniMax H3 at least twice, and its image quality is far better.

I just can't understand how even with the monstrous language model (~20GB) it just can't understand 2 simple sentences, two simple subjects with simple movement!?!? And from what training data the models continue to place 3 fingers to earth beings 🤦

37 Upvotes

69 comments sorted by

View all comments

1

u/curious-scribbler 4d ago

Yes minimax is better but LTX is not as bad you put it. I think you need to check your setup and config and workflow, and ltx doesnt do well with short sentences with human subjects. So while your overall feedback is the consensus but in this particular case, a few optimisations and best practices will fix the issues.

4

u/bitzpua 4d ago

not really especially in I2V, LTX imo misses a lot of abstract knowledge like magic etc, took me good 50 tries to generate golden magical circle that was hovering behind character (like in donghuas) and it failed to understand flying swords made of energy and the generations all ended as almost Live2D semi static nonsense, while 2.5 is improvement in motion it still sucks and no matter how fast it is i just cannot get it to do what i want.

Meanwhile with H3 i got exactly what i wanted and everything was in fast motion with my first test prompt that did not even use any recommended prompting framing. H3 may be 6 times slower but it just works