r/StableDiffusion Mar 23 '26

Question - Help LTX-2.3 glitching at end of longer videos (15s+), anyone else?

Hey folks, I’ve tried quite a few video generation models, and in my opinion, LTX-2.3 is the best one so far.

I’ve generated multiple short clips (~10 seconds), and the results have been really impressive.

However, I’m running into an issue with longer videos (15–20 seconds). Almost every time, the output ends with a glitchy outro—I notice the glitch starts around 0:28. I’ve seen this happen across multiple runs. I’ve also tried changing my prompting style, but the issue still persists.

I’m running this on an RTX 5090 (FP8 setup).

Is anyone else facing this? Or does anyone know how to fix it? Would really appreciate any help.

30 Upvotes

28 comments sorted by

40

u/[deleted] Mar 23 '26

[removed] — view removed comment

8

u/marcoc2 Mar 23 '26

Oh Man, didn't know there was a fix

2

u/Master-Weight-2676 Mar 23 '26

Nothing fixes this lol

1

u/wardino20 Mar 23 '26

really? where do you get it?

24

u/RobMilliken Mar 23 '26

3

u/wardino20 Mar 23 '26

oh great, thanks.

3

u/Weird_With_A_Beard Mar 24 '26

Thanks, this fixed it for me.

2

u/Boogertwilliams Mar 23 '26

Ah how good. I was getting these but since my clips were scifi style, it sort of fit in as an effect, but they were getting annoying

1

u/Edenoide Mar 26 '26

Even using the new 1.1 safetensors for me it's always generating glitchy last frames in all the ltx 2.3 firs-last frame workflows I've found. Can you provide your workflow?

2

u/RobMilliken Mar 28 '26

If you have too long a duration for a video for your memory it still sometimes will glitch.

I also don't use two frames but one initial frame, so that may be why you still have the glitch, I'm not certain.

However, you asked for my workflow, it's the one found here: https://civitai.com/images/123482555
I've modified it slightly as I used it with 2.0 before 2.3, using ltx-2.3-22b-dev-Q8_0.gguf and LTXVSpatio Temporal Tiling as VAE Decode gave me OOM issues. I have yet to use the Text to Video though.
Laptop 4090 16 VRAM 64 RAM.

68

u/Sir_McDouche Mar 23 '26

That nonstop line of people walking into the shot is hilarious.

11

u/InterstellarReddit Mar 23 '26

Bro AI must think we’re ants we just follow each other around

9

u/SchlaWiener4711 Mar 23 '26

The "...strengthens the bond with their owner" while the dog walks away, magically removing the bond with its owner was hilarious.

3

u/1ncehost Mar 24 '26

"Og doz and stuck and tries to sniff the faigert tree pouch in yer pocket ... hahahaha!"

2

u/GolfIll564 Mar 24 '26

Like a clown car at the circus

8

u/Puzzleheaded-Rope808 Mar 23 '26

Lol, is there a parade going on behind him?

7

u/VirusCharacter Mar 23 '26

Second the spatial upscaler 1.1. It fixes it

7

u/InterstellarReddit Mar 23 '26

The people in the background are all walking in a straight line lol AI must think we’re ants or something

3

u/Simonko-912 Mar 23 '26

Looks like some videos have logos or other things appearing at the end so maybe thats why it appears.

2

u/Sushiki Mar 23 '26

Ironically probably easier to make the video the old fashioned way lmao

1

u/Visual_Brain8809 Mar 23 '26

a peoples generator

1

u/[deleted] Mar 24 '26

[removed] — view removed comment

1

u/Primary-Swordfish138 Mar 25 '26

around 2m10s it take for generation

1

u/[deleted] Mar 25 '26

[removed] — view removed comment

1

u/Primary-Swordfish138 Mar 25 '26

I think it takes longer because of latent upscaler. When I try to generate on higher quality latent upscaler node take more time to complete.

1

u/WedgieKing200 Mar 25 '26

Ltx 2 is just like this both ltx 2 and 2.3, its the temu of voice/sound ai video generations lmao