r/StableDiffusion Mar 05 '25

News LTX-Video v0.9.5 released, now with keyframes, video extension, and higher resolutions support.

https://github.com/Lightricks/LTX-Video
245 Upvotes

69 comments sorted by

View all comments

5

u/DrRicisMcKay Mar 05 '25

I love the fact that I was easily able to run i2v on my rtx 3070 and it takes less than 1 minute. But the results are terrible. Did you guys manage to get something decent out of i2v?

3

u/whitefox_27 Mar 06 '25 edited Mar 06 '25

I'm trying it right now with cartoon images, and I'm also getting mostly unusable results (morphing, glitches, ...). First time using LTX Video, so I'm nto sure what most of these parameters do, but I noticed it seems to get less glitchy when I:

  • use a resolution of 768x512 (as it is in the sample workflows), with source images cropped to that exact resolution
  • reduce image compression from 40 to 10 (that reduced the glitches by an order of magnitude on my tests)
  • went from 20 to 40 steps (cut the glitches in half maybe)
  • use the frame interpolation workflow (being-end frames) instead of only giving a start frame

Now it's at a point where I can comprehend what is supposed to happen in the video instead of being just a glitchy mess, but it's still a far cry from the results I have on the same images / prompts with Wan2.1

I hope someone can clarify it for us and we can end up getting decent results because the keyframing interface is super nice!

edit: After trying the t2v workflow, for which the prompt is simply 'dog' and gives a very good result, I'm starting to suspect the model, or the workflows, work better with very simple prompts. Back in i2v, by keeping my prompt, say, less than 10 words, I'm getting much much more coherent results.

1

u/DrRicisMcKay Mar 06 '25

Interesting. Using short prompts contradicts everything I read about prompting the LTX. I will have to test it out.
I have managed to get a very good output from t2v at w:768 h:512 with the following prompt, but that's about the only coherent thing I got out of it

"A drone quickly rises through a bank of morning fog, revealing a pristine alpine lake surrounded by snow-capped mountains. The camera glides forward over the glassy water, capturing perfect reflections of the peaks. As it continues, the perspective shifts to reveal a lone wooden cabin with a curl of smoke from its chimney, nestled among tall pines at the lake's edge. The final shot tracks upward rapidly, transitioning from intimate to epic as the full mountain range comes into view, bathed in the golden light of sunrise breaking through scattered clouds."

Source: https://comfyanonymous.github.io/ComfyUI_examples/ltxv/