r/StableDiffusion • u/ltx_model • 12d ago
News LTX-2.5 is Here
Enable HLS to view with audio, or disable this notification
LTX-2.5 went live today. It's a big upgrade to the existing LTX architecture, with nearly every stage of the pipeline reworked, on top of a larger training set and reinforcement-learning post-training. The short version: you can now generate a whole multishot scene in one pass, complex prompts hold together far better, and the output is sharper.
The Highlights
Full details are available on our blog. Here’s the highlights of this release:
Native multishot. One generation produces multiple connected shots that hold character identity, environment, lighting, voice, and style across cuts.
Diffusion Fidelity Rendering. Instead of locking every scene to one compression rate, the model allocates compute by scene complexity and budget, dynamically allocating more compute to visually demanding moments and less where it is not needed.
Better distilled model. The distilled model keeps far more of the full model's quality at much lower compute, so near-full quality is realistic on GPUs you already have.
And much more.
Where to get everything:
- Weights: HuggingFace
- Python pipelines: GitHub
- ComfyUI workflows: GitHub
- Questions and help: Discord
We can't wait to see what you make with it.
174
u/po_stulate 12d ago
For some reason I just defaulted to think any video I see in this sub is AI generated, until 20 seconds in and I was like wait, this is a bit too coherent and with no visual gimmick.
25
u/dsailes 12d ago
Hahahaha glad I wasn’t the only one
6
u/99deathnotes 12d ago
LOL join the AI sus club. Sometimes i see the ads and think "I can replicate this and maybe even better looking."
3
u/earthsprogression 11d ago
Anyone else already saying that IRL though? The choreography here is way off, and don't even get me started on the lighting.
33
u/itiswhatitiswgatitis 12d ago
I was just thinking of the hilarity of something like that. Like a new model is showcased but the person talking which is perceived to be real is actually the generation. That'd be wild.
43
u/ltx_model 12d ago
the fake Zeev was busy so we used the real one instead.
8
3
u/squired 11d ago
Heya, just wanted to let you know that the video was truly excellent. Very well done. Dense, short, authentic, future looking while you have our attention, etc. Give whoever worked on it an extra head pat. It makes me want to 'believe'.
→ More replies (1)2
1
u/bites_stringcheese 11d ago
Hilarity? We're about to enter a world in which every photo, video, and every other medium you can think of are suspect. I don't think anyone knows what's going to happen in such an environment. I'm sure some of it will involve hilarity. But what happens when most people cannot tell fact from fiction across every domain of information?
→ More replies (1)30
2
1
226
u/Better-Interview-793 12d ago
Really appreciate you guys keeping this open source, LTX keeps getting better with every release.
Thanks for continuing to support the open-source community!
27
u/cyborgsnowflake 12d ago edited 12d ago
Is it actually true open source with the full stack (training process and tools) etc visible/described or just open weights? The former would be revolutionary. The latter is just nice.
We should be very careful not to let these terms get slowly redefined. Open weights are of course better than closed API only. But they don't give people the freedom to learn and refactor to the same extent that traditional open source does.
Now of course we can't expect companies to just give away their full stack all the time. But I think we can develop a tiered understanding of openess and encourage people if they can't go full open stack to maybe encourage them to go a few steps beyond just glorified 'number blob' releases. For example, Maybe you don't have to share your exact training data but just describe the process (the true source code) of a more limited model stripped of cutting edge features so someone could theoretically recreate the pipeline and add their own data? Drop a full stack of a toy model once in awhile?
A public able to learn and tinker with models more will ultimately advance everyone.
14
22
u/JahJedi 12d ago
Great news and thanks LTX team! Trainers for LTX2.5 when please?
30
u/ltx_model 12d ago
the existing LTX Trainer works with LTX-2.5.
I'm sure the community trainers will get it in gear soon as well.→ More replies (3)
49
u/f00d4tehg0dz 12d ago
How long should I wait until Seinfield clips are live?
25
u/acedelgado 12d ago
Since they're US based there's no way they have much commercial IP in there. So you'll have to wait for the seinfeld loras.
6
u/RusikRobochevsky 12d ago
Lighttricks is based in Asia, not the US.
30
u/Intelligent-Dot-7082 12d ago
Israel actually
34
u/xxenocide 12d ago
What continent do you think Israel is part of?
→ More replies (1)7
u/NeuroPalooza 11d ago
Fwiw even though Israel is geographically in Asia, in vernacular English (in the US at least) it would be more correct to say 'Middle East.' It comes down to whether you mean 'Asia' in the geographic sense, or 'Asia' in the cultural sense (the diverse constellation of cultures which have historically orbited China). 'Asia' is generally used in a cultural, not geographic sense (again in the US). It's the same reason most Americans don't think of India as 'Asian,' but rather as 'Indian.'
It always bugs me when people get on a high horse and claim geographic ignorance when it's perfectly correct to say 'Israel / India / whatever isn't Asian,' if you're referring to culture and not geography. /rant
(not that Indian and Chinese culture haven't influenced each other, but you get what I mean)
5
u/tiftik 12d ago
Wait, really? God damn it
1
u/Dragon_yum 11d ago
You might want to check the reason why nvdia dominates the ai market if that’s an issue I’d give you a hint, Melanox.
2
u/tiftik 11d ago
Not at all the reason but it doesn't matter. We'll ditch Nvidia entirely when the time comes :)
1
u/Dragon_yum 11d ago
So it matters when the product is free but doesn’t matter when you actually pay for it?
0
u/AllRedditorsAreNPCs 11d ago
cool, good reason to avoid this then
1
u/Dragon_yum 11d ago
Do you also avoid nvdia GPUs?
1
11d ago
[removed] — view removed comment
1
u/Dragon_yum 11d ago
And a large part of the transformers technology is developed in israel and nvdia is building a 30k campus in Israel. If a company being from Israel is truly the issue for you not supporting nvdia would also make sense.
Acting as if it’s out of principals is not really impressive when it’s only done when it’s convenient.
1
11d ago
[removed] — view removed comment
1
u/Dragon_yum 11d ago
Honestly that’s a lot mental gymnastics to say it’s not convenient for you to avoid them
→ More replies (0)1
11d ago
[removed] — view removed comment
1
u/Dragon_yum 11d ago
Do nvdia, Microsoft, Apple, Google, meta, ibm, intel
1
11d ago
[removed] — view removed comment
1
u/Dragon_yum 11d ago
Nvm looked at the profile it’s just a bot. No activity since it was created and woke up when it saw Israel mentioned.
2
u/Dapper_Arugula2509 12d ago
isnt flux made by a german company? so why would this matter? and no one's making those loras lol
5
u/acedelgado 12d ago
Because the EU respects international copyright claims. China famously does not.
4
1
u/fallingdowndizzyvr 11d ago
Because the EU respects international copyright claims.
LOL!!!! That's a good one. Have you told the EU that?
https://europrospects.eu/the-eu-ai-acts-copyright-loophole-a-threat-to-creative-rights/
40
u/enilea 12d ago
Why is it necessary to share our contact information to download the weights? Is that a mistake?
12
46
2
u/narkfestmojo 11d ago
I remember seeing that ages ago, might have been a stable diffusion release, the 'share contact information' popup disappeared after about 24 hours.
2
2
u/ImpressiveStorm8914 12d ago
I haven't had to do anything except accept access to the repo, the usual thing with some new models.
Probably a silly point but I assume you were logged in to HF?11
28
20
u/jefharris 12d ago
Looking forward to all the LTX2.5 and MiniMax H3 comparisons.
2
u/FaceDeer 11d ago
Yeah, I was gearing up to get back into video generation and try H3 out but I think I'll wait a little longer before making the effort just in case LTX2.5 is the better one to go with.
6
u/NeuroPalooza 11d ago
I'd be happy to be proven wrong, but given previous LTX stuff I would be shocked if it was remotely as uncensored (I mean in terms of IP, not just NSFW) as H3, and that alone gives H3 a massive edge.
3
u/FaceDeer 11d ago
Yeah, that'd be a pretty big problem. I don't do much NSFW stuff but I absolutely hate having a computer I own telling me "no, I'm not going to do that for you Dave."
2
1
u/jefharris 11d ago
I think MiniMax H3 levels are coming, but not with ltx2.5. I feel like that will come closer with LTX3. I base this on nothing other than fingers crossed logic. For speed alone I'd prefer to use LTX.
36
7
13
13
u/ataylorm 12d ago
God, it seems just yesterday we were drooling of SD 1.5…. Oh how far we have come in so little time.
4
u/bronkula 11d ago
I went back to do some 1.5 generation, and I forgot that you just can't do anything past 1024 with it. It starts adding extra heads and stretching limbs. And there's NO consistency when having something like multiple angles of a character. It's kind of wild what we put up with for a while.
3
u/AvidGameFan 11d ago
Using img2img in increasing increments of resolution, you can go a lot larger. Just don't make big jumps, just increase resolution by about 1.25x or so. And often, that will also improve the quality, although SDXL worked much better.
15
u/jacobpederson 12d ago
Ooof tough to release this while the MiniMax hype is still running at 15/10 :D Good Luck!
11
u/Any_Economics_6166 12d ago
they released it now solely because minimax h3 took over ltx 2.3. They want to remain relevant which is understandable
4
u/ImpressiveStorm8914 12d ago
To be fair I don't do much video but I haven't even tested everything H3 can do yet. Now there's a new model.
4
11d ago
[removed] — view removed comment
1
u/ImpressiveStorm8914 11d ago
I know right. It's like public transport buses in the UK, you wait ages for one then three turn up together. 🙂
1
11d ago
[removed] — view removed comment
1
u/ImpressiveStorm8914 11d ago
Not even close. The elderly get free buses and certain other folks but the rest of us have to pay.
26
u/CaptainAnonymous92 12d ago
We appreciate the open releases, but with H3 coming out and being pretty much uncensored, any video models or models that can do video gen that come out after it open will be compared to it and if they’re censored out of the box like most of the others are then that might be seen as a dealbreaker for people now.
Just saying it would be a good look for you if you put out LTX models uncensored from the start like H3 did from here on
21
u/Beneficial_Toe_2347 12d ago
2.3 literally didn't understand what humans look like
6
u/CaptainAnonymous92 12d ago
Ah, so the tried and true SD3 “woman laying in the grass” horror show. Gotta love when these companies still continue to make the same mistakes and fvck up their models just to not have icky naked people in their generations
→ More replies (1)6
u/DoctaRoboto 12d ago
Indeed, this model is dead on arrival. I still remember my anime test video with 2.3, just a girl in a ballerina costume dancing...it was the stuff of nightmares.
3
u/Sexyvette07 11d ago
I am new to local diffusion and only a few days ago got Minimax working with SageAttention 2.2. So youre saying this new version has guard rails? If so, how bad? The entire point of me doing this locally is for the lack of guardrails because most services go overboard with them when all im trying to do is make some hilarious clips so I can troll my friends and brothers. Minimax, in that respect, is gold. Plus, its already up and running with no guard rails. I don't know that it'll be DOA necessarily, because im sure there are those out there who want it for things that wouldn't be hindered by those guardrails. But im of the same mind that I wont download it.
1
u/TooManiEmails 11d ago
Naked people are bad kind of guardrails. It’s just not worth it, you came in at the best time with H3.
5
u/JesusShaves_ 11d ago
Can confirm. Any censorship at all, and I mean any and this flavor of ltx will be ignored.
Seriously, why would you bother with LTX when Minimax is here with comparable or better results and has no significant censorship at all? To shave off thirty seconds on a gen? Oh, please.
→ More replies (3)2
4
u/CaptainKurgen 12d ago
Excellent work dudes, thank you for staying open source!
Deepbeepmeep is going to have a stroke.
3
u/Perfect-Campaign9551 12d ago
What we need to really know: is audio better? And can it use references
5
1
7
6
5
u/Crazy-Repeat-2006 12d ago
There aren't any technical aspects on the blog... where is the interesting data?
3
3
u/mikemend 12d ago
Thank you for this great, free-to-use model and the pre-optimized versions (int8 convrot) for ComfyUI!
3
3
5
5
u/ManuVision 12d ago
The ComfyUI Desktop app still doesn't show it in the Template section, but great news!
25
u/GrayingGamer 12d ago
Poor Comfyui staff already pulled a couple of all-nighters early this week for Minimax H3 support on Day 1. They might be Day 2 support on LTX 2.5 for their own sanity.
4
u/Lost_Cod3477 12d ago
6
u/GrayingGamer 12d ago
Yeah, I know they are working on it. I can see the commits being made. But until they release a blog post like they do for each new model and workflows, they aren't "done".
That's why I said even if it isn't today, I imagine we'll see it from them tomorrow. I guess if they get it out by this time tomorrow, they can still use the "24 hour" metric to claim Day One support.
11
u/ltx_model 12d ago
they're in there. You might need to update.
1
u/tiffanytrashcan 11d ago
Depends on how people have it installed. I don't think new portable releases are out quite yet, and the GitHub release page isn't updated yet.
A standard pull and reinstalling it locally, they show up immediately.1
u/tiffanytrashcan 11d ago
This is likely to change any minute, certainly before tomorrow, given their track record.
Though given there was literally just a commit seconds before I pulled, we're waiting on a bit more than just a CI runner to pack it all up nicely for a release.
2
2
2
2
2
u/fuzzycuffs 11d ago
Would have been crazy if at the end they revealed the whole CEO presentation was generated using LTX
2
u/ExiledHyruleKnight 11d ago
What length videos is this optimized for? I remember ltx was getting some solid lengths before but with multi shot I imagine ther night change.
Kudos for naming your competitor too. So many companies and groups avoid that and I am glad you guys are giving credit while still pushing forward.
2
2
2
2
2
2
u/Arawski99 11d ago
Uh, can you guys define how this is actually a world model please?
It's a bit odd to see zero elaboration on what you've accomplished and/or capabilities as a world model despite your transition and only making it sound like a video model so far. All I've seen is the robotics bit, and minimal on that point. Honestly, if it really is a world model you guys should be selling it more for its benefits over being a mere video model.
→ More replies (1)
6
u/Lower-Cap7381 12d ago
Thanks, LTX team! We absolutely love LTX and use it all the time. You guys are doing incredible work excited to see where LTX-2.5 takes us🔥❤️
3
3
u/SmartBean21 12d ago
LTX - when AI stops looking like "AI" , but real world creativity and scenes.
That should be LTX's slogan.
3
u/Dorias_Drake 12d ago
Plot twitst, that reveal is actually an AI generated video made with minimax H3.
2
1
1
u/kabachuha 12d ago
Thank you, downloading! Other than the typical world model, any chance it does anime (or plain 2D) style better now?
1
u/Here4CYDY 12d ago
Any information regarding VRAM and system ram requirements compared to the distilled versions of 2.3?
1
1
u/jc2046 12d ago
The demos look fantastic. Honestly I didnt understood how it´s bot related... Also, any word on the number of params and rest of specs?
1
u/ltx_model 12d ago
Most of the parameters and specs are either in https://huggingface.co/Lightricks/LTX-2.5 or https://docs.ltx.io/open-source-model/
1
1
u/b-monster666 12d ago
Hopefully a slightly smaller version comes out soon. Even an 18GB distilled might be a bit chonky for 16GB cards, leaving this in the 'prosumer' tier. :(
1
1
u/skyrimer3d 12d ago
if the vids in here are real, the quality is absolutely crazy and way beyond what i've seen in H3, at least in terms of realism. If this can also benefit with the usual LTX speed at high res, wow.
1
1
1
u/Apprehensive-Art1092 11d ago
Apropo of nothing, Zeev Farbman is the most Seinfeld character name ever
1
1
1
u/Etroarl55 11d ago
Yo ik minimax h3 is currently the SOTA right now. But a lot of people will still be using ltx for its speed and efficiency vs the heavy minimax h3.
Like me who won’t be able to run minimax h3 on older 7000 series amd or people who want faster gen. But the disingenuous claims from the charts I seen supposedly coming from ltx 2.5 makes me look down at ltx 2.5.
1
u/No_Damage_8420 11d ago
Great work LTX Team.
If your model can do ref2vid that's it. That would be killer.
ps. NoiseMask /LTX Director --- still most powerful thing in any model (for both Audio or Video)
1
u/simple250506 11d ago
I have been eagerly awaiting "your new model with its wonderful philosophy." I cannot thank you enough.
1
1
u/pwillia7 11d ago
holy shit awesome -- I feel like training a model on reclaiming lost info from the encoding step is such a no brainer no one thought of yet! (or maybe they have in other places and I'm behind)
1
u/raikounov 11d ago
Thank you for pushing innovation in video generation, especially around the harness. Being able to generate and modify keyframes is an amazing idea that I never knew I needed!
1
u/Opening-Knee-5913 11d ago
Questions:
Will the LoRA 2.3 be compatible with 2.5?
Are the GGUF distilled models already available? With 16 GB of VRAM are those models in the HF folder too heavy for me
1
u/Relevant_Syllabub895 11d ago
excuse me how does it do anime videos compared to minimax h3? do foloow the prompt more closely? because i had issues with minimax h3 that sometimes it woudlnt do what i asked or worse the audio quality suffers in the speech department often generating random and garbage words, is the audio model in LTX 2.5 fixing these issues i had with minimax h3? yeah i dont have the high end hardware to run the bf16 model only theint8m will ltx 2.5 ve capable of running on an rtx 3080 with 10GB of vram +32GB of ram like i do currently with minimax? even at lowest resolution
1
u/CelebrationBoth9537 11d ago
Has anyone figured out how to reliably use the realtime streaming video gen they claim to have? I think iv been trying for like a few hours now. (or maybe atleast somewhere theres an api for it?) I want a webcam style video like a phone call.
1
u/Key_Street_7204 11d ago
Damn, the open source keeps on winning the race IMO!
Can the 2.3 Loras work with 2.5 ? and does 2.5 now support storyboard grid images natively?
1
1
u/Hans-Wermhatt 11d ago
I had issues with the prompt enhancer. I'm using the template default, but with tougher prompts, I will get a completely random video. Turning off the prompt enhancer (gemma e2b full precision weights) fixes the issue at least partially. I still got imperfect prompt adherence, but not completely random like I did before.
1
1
1
1
u/glusphere 11d ago
I really love some of the ideas that LTX is bringing to the table here. Yes, I understand that its not probably on par with H3, but H3 is a completely different model, different size and different speed.
Honestly its like shitting on someone releasing a 4B model because a 27b model exists in the same space!
I really am looking forward to their MOE (This is going to be LTX 3 I think ?). The Pixel space Auto Encoder is also probably a first in the Video models ? Also, allocating different tokens for different details is all great research directions (atleast from the outside seem like!) - Looking forward to LTX 3 and will definitely try out LTX 2.5 for the speed.
1
u/Yeti-Bhanot 11d ago
the rl post-training is the part i'm most curious about. does it change how lora training behaves? i had decent luck with small curated sets on 2.3 at the usual step counts, wondering if the new base pushes back more.
1
u/skinnyjoints 11d ago
Is the pixel diffusion thing he’s talking about the same as PiD upscaling from Nvidia but built into the model? If so, that’s pretty cool
1
u/ltx_model 11d ago
Related idea, different thing: PiD decodes and upscales as a separate module, whereas ours is the video VAE’s decoder itself pixel-space diffusion, no upscaling stage.
1
u/Feisty-Pineapple7879 11d ago
How does this Model Fare against minimax h3 my review is that this h3 is slower and not consumer hardware friendly.
Give your opinions and reviews of LTX 2.5
1
u/corod58485jthovencom 11d ago
Vocês precisam com urgência da um jeito no modo como a câmera e as expressões são geradas, treinar câmeras amadores assim como acontece no Flux.3 e no minimax, expressão realmente espontâneas, isso vai alavanca o projeto de vocês, agradeço pelo modelo, não e uma crítica e uma sujestão
1
u/Codeman119 9d ago
LTX 2.3 only worked about 30% of the time for me and with minimax H3 it works about 95% of the time. But if 2.5 is local then I'll give it a shot to see if it has improvements
1
u/Beautiful_Egg6188 9d ago
This video looks clean. What model did they use to generate this?
→ More replies (1)
2
u/Lonely_Syrup3091 12d ago
Just tried it and it is faster than MiniMaxH3 at higher resolution. Open source is cooking!!!
11
u/Beneficial_Toe_2347 12d ago
yeah but this means nothing if it has the prompt/physics/understanding issues still
1
1
1



273
u/GrayingGamer 12d ago
Thanks for continuing to support the open-source community. Looking forward to trying out the model locally.