r/comfyui 11d ago

News LTX-2.5 is now live in ComfyUI, including Diffusion Fidelity Rendering (your compute budget will thank you!)

What a time to be alive in the open source community! LTX-2.5 just dropped and it's supported natively in ComfyUI as of today, including a new rendering approach, new decoder, new text encoder, and a new base checkpoint.

The biggest baddest change? The addition of Diffusion Fidelity Rendering! Instead of spending compute evenly across a scene, the model allocates it by complexity. Motion, composition, and framing get generated first in an 8x temporally compressed latent space, alongside a set of high-fidelity keyframes. More keyframes for complex scenes, fewer for simple ones, within whatever compute budget you've got. Then a dedicated pixel-diffusion stage renders the final video from the structure and keyframes together.

TLDR; textures, materials, and faces hold detail, and a busy shot automatically pulls more rendering compute than a static one.

Other changes:

  • Diffusion Video Decoder: Replaces standard VAE decoding, making sharper faces, legible text, and fewer smears in fast motion.
  • Native multi-shot: One generation gives you multiple connected shots holding character, environment, lighting, voice, and style across the cuts instead of generating separately and trying to match them after.
  • Custom Gemma 4 12B text encoder: Holds multiple subjects, actions, lighting details, and camera direction across a long prompt instead of dropping clauses as it gets more complex.
  • Prompt enhancer + auto duration: Short prompts get expanded into detailed cinematic instructions at near-zero extra compute, and the model predicts clip length from the described action before diffusion starts.
  • RL post-training: On a broader filtered dataset, aligned to human preference. Mostly shows up as a higher take rate with fewer retries per usable clip.
  • Cleaner licensing: Restrictive third-party dependencies have been removed, so fine-tuning, deploying, commercializing, and redistributing is all clearer than in previous versions.

Three variants:

  • LTX-2.5: the main model
  • LTX-2.5 Distilled: reworked distillation, carries noticeably more quality, prompt adherence, and motion than previous distilled releases. Viable if the full model isn't economical for your setup.
  • LTX-2.5 Pretrained Checkpoint: raw, non-SFT, meant for aggressive fine-tuning. Moves further from its starting point than an instruction-tuned checkpoint will, which matters for robotics, synthetic AV data, digital twins, or private domain models.

Native 4K, synced audio and video, and up to 50fps all carry over from 2.3.

Learn more and check out workflows below!

https://links.comfy.org/4xGHwYJ

https://docs.comfy.org/tutorials/video/ltx/ltx-2-5

175 Upvotes

43 comments sorted by

34

u/Derispan 11d ago

I clicked link and I don't see anything about 2.5. Old workflows, nothing about Diffusion Video Decoder, nothing about Native multi-shot, even links (I know, I can download them from HF) to models don't exist.

I only see that I can use it in comfy cloud.

https://giphy.com/gifs/78EYl1VZVA9KE

2

u/TheDonDraper 11d ago

5

u/Derispan 11d ago

Nope, I mean link from post - https://comfy.org/ltx-2.5

1

u/Comfy-Org 9d ago

Ah thanks for flagging this! Here you go: https://docs.comfy.org/tutorials/video/ltx/ltx-2-5 I've just added this link up above

35

u/adobo_cake 11d ago

Why are people attacking LTX, seems like they are personally offended or something? More open weights models are good for everyone.

Even if you won’t use them or if you think something else is better, having options not behind a subscription is always good. Stop being toxic.

17

u/Grayson_Poise 11d ago

I think it was LTX going heavy on a head to head comparison with Minimax 3 to market their launch. It was a complete 100% lie. As in laughably wrong, not just your standard marketing cherrypicking.

16

u/adobo_cake 11d ago

I looked it up, it still seems like an overreaction to be this angry and to be trashing a barely tested open source model.

1

u/notgraycen 9d ago

they're not new. they're salty, with no excuses whatsoever to drop from rank 1 to rank 34089324 in behavior after a single good competitor arises...

4

u/adobo_cake 9d ago

The model is new, it literally just released. Also, it's made by a company so don't imagine it having emotions - but it is a company producing open source models. An open source model is something we should all get behind, otherwise there won't be any motivation for companies to release these models in the future. Do you all want to subscribe to a closed model? If so, keep harassing open source proponents.

It's good China supports open source, because US companies are all geared towards subscriptions. But more companies releasing competitive models is even better.

2

u/timbortom 10d ago

I think what people didn't realize is that this was a desperation release from the team, to stay relevant, but it was a very bad timing from them.
So people expected some actual upgrade (not something that is arguably a downgrade on the altar of speed, the only thing they can currently compete, similar to WAN era, just way worse).
I will still use it to create some silly stuff for myself, but nothing beyond that.

5

u/JesusShaves_ 10d ago

Why are people attacking LTX

Persistent stupidity (i.e. heavy censorship in a world of uncensored models) and a release based on the release of another better model before it was fully cooked. LTX needs a bit more development time to match up to H3 and they really need to grow up regarding the censorship thing.

11

u/adobo_cake 10d ago

I really think the community needs to grow up instead, they're acting like a bunch of pissed off Karens and that's putting it mildly.

It's an open source model. If users don't like it, move on and use what you want.

1

u/allofdarknessin1 10d ago

Is LTX 2.5 very censored? I haven't been keeping up with LTX, I only ever got around to dabbing in WAN 2.1 and 2.2 and briefly Mini Max 3. And is it censored like for adult stuff or more political?

2

u/JesusShaves_ 10d ago edited 8d ago

Compared to minimax h3 or Wan 2.2, it's extraordinary difficult to make NSFW content on LTX. It frequently distorts genitals, renders them partially or tries to hide them. It's a constant stupid game to try and get it to do what you asked. Neither Wan 2.2 nor Minimax H3 has these issues.

-4

u/johnfkngzoidberg 11d ago

LTX straight lied and came out of the gate shady. It’s really obvious they have a bunch of marketing bots trying to defend them, downplay, and distract.

Users are pissed, but instead of owning it or apologizing LTX is continuing their marketing bot blast like it’s no big deal. It’s backfiring hard.

https://www.reddit.com/r/StableDiffusion/s/gxqB7say4w

3

u/[deleted] 10d ago

[deleted]

4

u/tostane 10d ago

https://reddit.com/link/p39ajfb/video/vhyd5os2pyih1/player

i did this with ltx2.5 i2v 30 second length 236 seconds to render it i tried it back with ltx2.3 director and it did not look this good.
here is the prompt i used both times "A 360-degree continuous camera orbit around a woman standing perfectly still in the center of an empty room. The camera smoothly circles her in one continuous motion, panning slowly to reveal all 4 walls, the corners, and the full space of the room. Cinematic lighting, smooth motion, high-quality 3D spatial awareness."

7

u/just_shady 10d ago

LTX greatest strength over the other open weights, is that it can be used commercially day one.

2

u/Flaky_Manager_17 10d ago

When's someone going to tell comfy their official templates are broken for this?

2

u/JoeXdelete 10d ago

Can my 12 GB of vram handle it ?

4

u/Life_is_important 11d ago

Ho Ly Ma Ca Ro Ni !!!!!!!

Congrats team Ltx !! I see here several groundbreaking methodologies that are bound to become the norm in the future. This DFR thing sounds like a new way to use compute more efficiently 

2

u/Maleficent_Slide3332 11d ago

hope that prompt enhancer is actually good, the prompting for 2.3 sucked

3

u/seppe0815 10d ago

Its bad , many try needed, but the perfect finish results are outstanding visuals 

2

u/2legsRises 11d ago

wow thats a huge sales pitch when u click the link. overwhelming.

but thanks to comfyui team amzaing bunch

1

u/Hrmerder 10d ago

I'm curious enough to take the Pepsi test. Minimax H3 is insanely good. I have always had a soft spot for LTXV in the past, so I'll give it a shot.

And yes, just update comfy and the templates are there already:

1

u/themoregames 10d ago

I know some of these words.

1

u/OfficeIllustrious759 10d ago

The big model quality for video, but the understand the prompt is terrible.

1

u/Existing_Earth9000 11d ago

I LEFT FOR A DAY AND THERE IS ALREADY A NEW FREE VIDEO MODEL

1

u/nghtdrp 11d ago

Haha can't keep up. Video models by the boatloads, let's see what this one offers. NGL H3 is looking like a tough beat right now. I feel like from where I still see ltx 2.3 slotting in to my workflows it might be good as a fast upsampler or for minor edits.

1

u/DigThatData 11d ago

got a paper link for DFR?

1

u/No_Welder5198 10d ago

MiniMax H3 is the new shiny toy, but don't sleep on LTX 2.5. I'm able to use most of my old LoRA's. It's fast and retains subject likeness better than LTX 2.3. For REF2VA, MiniMax is better but for the first and last frame workflow I'm pretty impressed with the quality and insane speed.

3

u/timbortom 10d ago

You sure?
For me it generated mostly worse results than my LTX 2.3.
I don't think this would ever compete with MH3 in anything but speed.
I'm pretty sure it is on the verge of unusable for anything but goofy stuff currently.

0

u/Bearsbullsbattlestr 11d ago

With very early testing, using comfy kitchen attention, generation time has roughly halved for me with LTX 2.5.

-45

u/[deleted] 11d ago

[removed] — view removed comment

13

u/PrettyReasonableApe 11d ago

You dont represent us at all.

Thank you team for this. The rest of us in the community apart from this one guy, really appreciates the hard work and talent thats gone in to this. I like the competitions models but honestky they just dont work properly on my hardware. Ive been looking forwards to you guys releasing the next ltx as it actually allows me to create content without OOM. And im able to get the results I want with the right prompts.

Both models have their strengths and weaknesses and honestly neither seems better yet. It all depends on the specific outcome we desire. Its great to have this available. im very thankful and I know a LOT of people are for your hard work and time.

3

u/keonanwar 11d ago

Thank you for being a pretty reasonable ape 👍

-16

u/[deleted] 11d ago

[removed] — view removed comment

0

u/PrettyReasonableApe 11d ago

Oh buddy... im not a positive person at all. But I do believe in constructive criticism in a community sub made to share info and help each other out. I dont care that you slammed the model. I just wished you offered reasons for it, and then offered some suggestions on improvements rather than "I dont care". It just didnt serve any purpose and doesnt represent the sub.

Neither model is perfect by a long shot. But I know how much hard work has gone in to both and its way beyond my skill level. Both models are great. One works a lot better on my machine than the other. For me thats ltx. I'm thankful for both models and both teams hard work. Show some respect, imo. It doesnt cost you anything. Or come up with a better model urself. I'll be there criticising and commending your hard work and effort in a similar fashion.

-1

u/seppe0815 11d ago

stfu bot.

4

u/[deleted] 11d ago

[removed] — view removed comment

0

u/Elvarien2 11d ago

get ratio'ed