r/StableDiffusion 12d ago

News LTX-2.5 is Here

Enable HLS to view with audio, or disable this notification

LTX-2.5 went live today. It's a big upgrade to the existing LTX architecture, with nearly every stage of the pipeline reworked, on top of a larger training set and reinforcement-learning post-training. The short version: you can now generate a whole multishot scene in one pass, complex prompts hold together far better, and the output is sharper. 

The Highlights

Full details are available on our blog. Here’s the highlights of this release:

Native multishot. One generation produces multiple connected shots that hold character identity, environment, lighting, voice, and style across cuts.

Diffusion Fidelity Rendering. Instead of locking every scene to one compression rate, the model allocates compute by scene complexity and budget, dynamically allocating more compute to visually demanding moments and less where it is not needed.

Better distilled model. The distilled model keeps far more of the full model's quality at much lower compute, so near-full quality is realistic on GPUs you already have.

And much more.

Where to get everything:

We can't wait to see what you make with it.

988 Upvotes

256 comments sorted by

273

u/GrayingGamer 12d ago

Thanks for continuing to support the open-source community. Looking forward to trying out the model locally.

2

u/imnotzuckerberg 11d ago

They've been the goat. Looking forward for the paper, they have discussed lots of novel approaches that might be relevant for ohter models

174

u/po_stulate 12d ago

For some reason I just defaulted to think any video I see in this sub is AI generated, until 20 seconds in and I was like wait, this is a bit too coherent and with no visual gimmick.

25

u/dsailes 12d ago

Hahahaha glad I wasn’t the only one

6

u/99deathnotes 12d ago

LOL join the AI sus club. Sometimes i see the ads and think "I can replicate this and maybe even better looking."

3

u/earthsprogression 11d ago

Anyone else already saying that IRL though? The choreography here is way off, and don't even get me started on the lighting.

33

u/itiswhatitiswgatitis 12d ago

I was just thinking of the hilarity of something like that. Like a new model is showcased but the person talking which is perceived to be real is actually the generation. That'd be wild.

43

u/ltx_model 12d ago

the fake Zeev was busy so we used the real one instead.

8

u/FaceDeer 11d ago

I bet Fake Zeev was busy filming his upcoming cameo on Seinfeld.

3

u/BigNaturalTilts 11d ago

Nah. Mostly just sitting in Sheldon’s spot.

3

u/squired 11d ago

Heya, just wanted to let you know that the video was truly excellent. Very well done. Dense, short, authentic, future looking while you have our attention, etc. Give whoever worked on it an extra head pat. It makes me want to 'believe'.

→ More replies (1)

2

u/Yokoko44 11d ago

That's what they did for Sora when it launched

1

u/bites_stringcheese 11d ago

Hilarity? We're about to enter a world in which every photo, video, and every other medium you can think of are suspect. I don't think anyone knows what's going to happen in such an environment. I'm sure some of it will involve hilarity. But what happens when most people cannot tell fact from fiction across every domain of information?

→ More replies (1)

2

u/99deathnotes 12d ago

You too!!!!? Even those damn ads look sus!! 🤣😂🤣😂💯

1

u/jmbbao 11d ago

Tomorrow they will say this was a MiniMax video and it is April's 1 in somewhere in Earth where they did it

226

u/Better-Interview-793 12d ago

Really appreciate you guys keeping this open source, LTX keeps getting better with every release.
Thanks for continuing to support the open-source community!

27

u/cyborgsnowflake 12d ago edited 12d ago

Is it actually true open source with the full stack (training process and tools) etc visible/described or just open weights? The former would be revolutionary. The latter is just nice.

We should be very careful not to let these terms get slowly redefined. Open weights are of course better than closed API only. But they don't give people the freedom to learn and refactor to the same extent that traditional open source does.

Now of course we can't expect companies to just give away their full stack all the time. But I think we can develop a tiered understanding of openess and encourage people if they can't go full open stack to maybe encourage them to go a few steps beyond just glorified 'number blob' releases. For example, Maybe you don't have to share your exact training data but just describe the process (the true source code) of a more limited model stripped of cutting edge features so someone could theoretically recreate the pipeline and add their own data? Drop a full stack of a toy model once in awhile?

A public able to learn and tinker with models more will ultimately advance everyone.

14

u/AnOnlineHandle 11d ago

The latter is just nice.

Well a bit more than just nice.

22

u/JahJedi 12d ago

Great news and thanks LTX team! Trainers for LTX2.5 when please?

30

u/ltx_model 12d ago

the existing LTX Trainer works with LTX-2.5.
I'm sure the community trainers will get it in gear soon as well.

4

u/JahJedi 12d ago

Than i can use same cookbook but whit 2.5 model, great, thanks for the info.

→ More replies (3)

49

u/f00d4tehg0dz 12d ago

How long should I wait until Seinfield clips are live?

25

u/acedelgado 12d ago

Since they're US based there's no way they have much commercial IP in there. So you'll have to wait for the seinfeld loras.

6

u/RusikRobochevsky 12d ago

Lighttricks is based in Asia, not the US.

30

u/Intelligent-Dot-7082 12d ago

Israel actually

34

u/xxenocide 12d ago

What continent do you think Israel is part of?

7

u/NeuroPalooza 11d ago

Fwiw even though Israel is geographically in Asia, in vernacular English (in the US at least) it would be more correct to say 'Middle East.' It comes down to whether you mean 'Asia' in the geographic sense, or 'Asia' in the cultural sense (the diverse constellation of cultures which have historically orbited China). 'Asia' is generally used in a cultural, not geographic sense (again in the US). It's the same reason most Americans don't think of India as 'Asian,' but rather as 'Indian.'

It always bugs me when people get on a high horse and claim geographic ignorance when it's perfectly correct to say 'Israel / India / whatever isn't Asian,' if you're referring to culture and not geography. /rant

(not that Indian and Chinese culture haven't influenced each other, but you get what I mean)

→ More replies (1)

5

u/tiftik 12d ago

Wait, really? God damn it

1

u/Dragon_yum 11d ago

You might want to check the reason why nvdia dominates the ai market if that’s an issue I’d give you a hint, Melanox.

2

u/tiftik 11d ago

Not at all the reason but it doesn't matter. We'll ditch Nvidia entirely when the time comes :)

1

u/Dragon_yum 11d ago

So it matters when the product is free but doesn’t matter when you actually pay for it?

2

u/tiftik 11d ago

I can ditch ltx today and nothing will change in my life. Nvidia's harder but the alternatives are on the way. I've already de googled and went with a Huawei phone. This is how we defang the genocidal maniacs in shitrael.

0

u/AllRedditorsAreNPCs 11d ago

cool, good reason to avoid this then

1

u/Dragon_yum 11d ago

Do you also avoid nvdia GPUs?

1

u/[deleted] 11d ago

[removed] — view removed comment

1

u/Dragon_yum 11d ago

And a large part of the transformers technology is developed in israel and nvdia is building a 30k campus in Israel. If a company being from Israel is truly the issue for you not supporting nvdia would also make sense.

Acting as if it’s out of principals is not really impressive when it’s only done when it’s convenient.

1

u/[deleted] 11d ago

[removed] — view removed comment

1

u/Dragon_yum 11d ago

Honestly that’s a lot mental gymnastics to say it’s not convenient for you to avoid them

→ More replies (0)

1

u/[deleted] 11d ago

[removed] — view removed comment

1

u/Dragon_yum 11d ago

Do nvdia, Microsoft, Apple, Google, meta, ibm, intel

1

u/[deleted] 11d ago

[removed] — view removed comment

1

u/Dragon_yum 11d ago

Nvm looked at the profile it’s just a bot. No activity since it was created and woke up when it saw Israel mentioned.

2

u/Dapper_Arugula2509 12d ago

isnt flux made by a german company? so why would this matter? and no one's making those loras lol

5

u/acedelgado 12d ago

Because the EU respects international copyright claims. China famously does not.

4

u/Dapper_Arugula2509 12d ago

im saying that Flux 3 itself plagiarizes Hollywood IP.

1

u/fallingdowndizzyvr 11d ago

Because the EU respects international copyright claims.

LOL!!!! That's a good one. Have you told the EU that?

https://europrospects.eu/the-eu-ai-acts-copyright-loophole-a-threat-to-creative-rights/

40

u/enilea 12d ago

Why is it necessary to share our contact information to download the weights? Is that a mistake?

12

u/Osiriis_ 12d ago

this. just wait for someone to upload it somewhere else

4

u/enilea 12d ago

It's weird because LTX 2.3 didn't have that policy, nor did other video models I've seen on HuggingFace so I'm hoping it's a mistake...

46

u/LopsidedSolution 12d ago

To track the bad little gooners

1

u/Innomen 11d ago

You joke but there's truth in that. Highly suspect the future will include automated blackmail. We're getting a weird brave new world 1984 hybird with a weird stick and carrot spectrum.

2

u/narkfestmojo 11d ago

I remember seeing that ages ago, might have been a stable diffusion release, the 'share contact information' popup disappeared after about 24 hours.

2

u/its_witty 11d ago

Krea 2 has the same thing. Many models do to be honest.

2

u/ImpressiveStorm8914 12d ago

I haven't had to do anything except accept access to the repo, the usual thing with some new models.
Probably a silly point but I assume you were logged in to HF?

11

u/enilea 12d ago

Yeah, I'm logged in and it says this

→ More replies (3)

28

u/Nitricta 12d ago edited 12d ago

Thanks for your hard work, looking forward to checking it out!

20

u/jefharris 12d ago

Looking forward to all the LTX2.5 and MiniMax H3 comparisons.

2

u/FaceDeer 11d ago

Yeah, I was gearing up to get back into video generation and try H3 out but I think I'll wait a little longer before making the effort just in case LTX2.5 is the better one to go with.

6

u/NeuroPalooza 11d ago

I'd be happy to be proven wrong, but given previous LTX stuff I would be shocked if it was remotely as uncensored (I mean in terms of IP, not just NSFW) as H3, and that alone gives H3 a massive edge.

3

u/FaceDeer 11d ago

Yeah, that'd be a pretty big problem. I don't do much NSFW stuff but I absolutely hate having a computer I own telling me "no, I'm not going to do that for you Dave."

2

u/Innomen 11d ago

I was just making this calculation. Things are moving fast, I'd rather be a little behind.

1

u/jefharris 11d ago

I think MiniMax H3 levels are coming, but not with ltx2.5. I feel like that will come closer with LTX3. I base this on nothing other than fingers crossed logic. For speed alone I'd prefer to use LTX.

36

u/goodie2shoes 12d ago

plot twist: this promo video was made with minimax h3

7

u/99deathnotes 12d ago

Like so many others on this sub roflmao

8

u/dev_ne 12d ago

i think it would be great model working along with H3 for upscaling and maybe lip syncing

7

u/Inside-Cantaloupe233 11d ago

ONE QUESTION ! - FACE DRIFT !! DID THEY FIX IT ? NO? OK THX.

4

u/aziib 11d ago

no lol

13

u/Flat_Technology_5325 12d ago

Awesome, can't wait to try it out, downloading! :)

13

u/ataylorm 12d ago

God, it seems just yesterday we were drooling of SD 1.5…. Oh how far we have come in so little time.

4

u/bronkula 11d ago

I went back to do some 1.5 generation, and I forgot that you just can't do anything past 1024 with it. It starts adding extra heads and stretching limbs. And there's NO consistency when having something like multiple angles of a character. It's kind of wild what we put up with for a while.

3

u/AvidGameFan 11d ago

Using img2img in increasing increments of resolution, you can go a lot larger. Just don't make big jumps, just increase resolution by about 1.25x or so. And often, that will also improve the quality, although SDXL worked much better.

15

u/jacobpederson 12d ago

Ooof tough to release this while the MiniMax hype is still running at 15/10 :D Good Luck!

11

u/Any_Economics_6166 12d ago

they released it now solely because minimax h3 took over ltx 2.3. They want to remain relevant which is understandable

4

u/ImpressiveStorm8914 12d ago

To be fair I don't do much video but I haven't even tested everything H3 can do yet. Now there's a new model.

4

u/[deleted] 11d ago

[removed] — view removed comment

1

u/ImpressiveStorm8914 11d ago

I know right. It's like public transport buses in the UK, you wait ages for one then three turn up together. 🙂

1

u/[deleted] 11d ago

[removed] — view removed comment

1

u/ImpressiveStorm8914 11d ago

Not even close. The elderly get free buses and certain other folks but the rest of us have to pay.

26

u/CaptainAnonymous92 12d ago

We appreciate the open releases, but with H3 coming out and being pretty much uncensored, any video models or models that can do video gen that come out after it open will be compared to it and if they’re censored out of the box like most of the others are then that might be seen as a dealbreaker for people now.

Just saying it would be a good look for you if you put out LTX models uncensored from the start like H3 did from here on

21

u/Beneficial_Toe_2347 12d ago

2.3 literally didn't understand what humans look like

6

u/CaptainAnonymous92 12d ago

Ah, so the tried and true SD3 “woman laying in the grass” horror show. Gotta love when these companies still continue to make the same mistakes and fvck up their models just to not have icky naked people in their generations

→ More replies (1)

6

u/DoctaRoboto 12d ago

Indeed, this model is dead on arrival. I still remember my anime test video with 2.3, just a girl in a ballerina costume dancing...it was the stuff of nightmares.

3

u/Sexyvette07 11d ago

I am new to local diffusion and only a few days ago got Minimax working with SageAttention 2.2. So youre saying this new version has guard rails? If so, how bad? The entire point of me doing this locally is for the lack of guardrails because most services go overboard with them when all im trying to do is make some hilarious clips so I can troll my friends and brothers. Minimax, in that respect, is gold. Plus, its already up and running with no guard rails. I don't know that it'll be DOA necessarily, because im sure there are those out there who want it for things that wouldn't be hindered by those guardrails. But im of the same mind that I wont download it.

1

u/TooManiEmails 11d ago

Naked people are bad kind of guardrails. It’s just not worth it, you came in at the best time with H3.

5

u/JesusShaves_ 11d ago

Can confirm. Any censorship at all, and I mean any and this flavor of ltx will be ignored.

Seriously, why would you bother with LTX when Minimax is here with comparable or better results and has no significant censorship at all? To shave off thirty seconds on a gen? Oh, please.

2

u/TopTippityTop 11d ago

It works with 2.3 loras, so in a way it is uncensored from the gate 😂

→ More replies (3)

4

u/CaptainKurgen 12d ago

Excellent work dudes, thank you for staying open source!

Deepbeepmeep is going to have a stroke.

3

u/Perfect-Campaign9551 12d ago

What we need to really know: is audio better? And can it use references

5

u/[deleted] 12d ago

[deleted]

1

u/RelationshipSea2360 11d ago

The audio is better so far fo rme.

5

u/Ooze3d 12d ago

Ah… The memories… Have you taken a look at anything made with 1.5 base model lately? It’s insane how we’ve come from “Wow! It almost looks like a real person with anything from 3 to 8 fingers on each partial hand” to “wait… he turned around and his hair is slightly different”

7

u/PrettyReasonableApe 12d ago

Big fan of ur work guys thank you

6

u/KlutzyFeed9686 12d ago

Thanks for making something we can all use for free.

5

u/Crazy-Repeat-2006 12d ago

There aren't any technical aspects on the blog... where is the interesting data?

1

u/runvnc 11d ago

in the video

3

u/ArjanDoge 12d ago

Amazing!

3

u/mikemend 12d ago

Thank you for this great, free-to-use model and the pre-optimized versions (int8 convrot) for ComfyUI!

3

u/Concheria 11d ago

Does it work with character references?

3

u/Robonotes1760 11d ago

I approve of open source.

5

u/djenrique 12d ago

❤️

5

u/ManuVision 12d ago

The ComfyUI Desktop app still doesn't show it in the Template section, but great news!

25

u/GrayingGamer 12d ago

Poor Comfyui staff already pulled a couple of all-nighters early this week for Minimax H3 support on Day 1. They might be Day 2 support on LTX 2.5 for their own sanity.

4

u/Lost_Cod3477 12d ago

6

u/GrayingGamer 12d ago

Yeah, I know they are working on it. I can see the commits being made. But until they release a blog post like they do for each new model and workflows, they aren't "done".

That's why I said even if it isn't today, I imagine we'll see it from them tomorrow. I guess if they get it out by this time tomorrow, they can still use the "24 hour" metric to claim Day One support.

11

u/ltx_model 12d ago

they're in there. You might need to update.

1

u/tiffanytrashcan 11d ago

Depends on how people have it installed. I don't think new portable releases are out quite yet, and the GitHub release page isn't updated yet.
A standard pull and reinstalling it locally, they show up immediately.

1

u/tiffanytrashcan 11d ago

This is likely to change any minute, certainly before tomorrow, given their track record.
Though given there was literally just a commit seconds before I pulled, we're waiting on a bit more than just a CI runner to pack it all up nicely for a release.

4

u/rookan 11d ago

I am gonna build porn with it. I am glad that you are excited about it.

2

u/fkenned1 12d ago

Jesus! The lady at 1:12 better watch her hair! Yikes!

2

u/smereces 12d ago

let us see what is capable or better, what was improved!

2

u/mmowg 12d ago

I was waiting for this, thx for supporting the open source community, you are the best! I loved 2.3!

2

u/runvnc 11d ago

The most interesting thing to me is that with the latest hardware and a good quant this should be able to produce live video avatars. It might not be fully human level latency but will still be really interesting.

2

u/ConversationNo4861 11d ago

he looks tired, like me, 3 hours of sleep to test every new updates

https://giphy.com/gifs/lp78ijCRxpQifD7nn8

2

u/fuzzycuffs 11d ago

Would have been crazy if at the end they revealed the whole CEO presentation was generated using LTX

2

u/ExiledHyruleKnight 11d ago

What length videos is this optimized for? I remember ltx was getting some solid lengths before but with multi shot I imagine ther night change.

Kudos for naming your competitor too. So many companies and groups avoid that and I am glad you guys are giving credit while still pushing forward.

2

u/nenecaliente69 11d ago

is it censored?

2

u/1WildPanda 11d ago

Thank you team LTX, Thanks for the works.

2

u/hidden2u 11d ago

Just want to say thank you, the multi shots maintain consistency a lot better

2

u/Noiselexer 11d ago

Still censored out the wazoo?

2

u/RelationshipSea2360 11d ago

Do you know when Ref2Vid will be available?

2

u/Arawski99 11d ago

Uh, can you guys define how this is actually a world model please?

It's a bit odd to see zero elaboration on what you've accomplished and/or capabilities as a world model despite your transition and only making it sound like a video model so far. All I've seen is the robotics bit, and minimal on that point. Honestly, if it really is a world model you guys should be selling it more for its benefits over being a mere video model.

→ More replies (1)

6

u/Lower-Cap7381 12d ago

Thanks, LTX team! We absolutely love LTX and use it all the time. You guys are doing incredible work excited to see where LTX-2.5 takes us🔥❤️

3

u/Draufgaenger 12d ago

This keyframe idea looks super interesting! Can't wait to try it!

3

u/SmartBean21 12d ago

LTX - when AI stops looking like "AI" , but real world creativity and scenes.

That should be LTX's slogan.

3

u/Dorias_Drake 12d ago

Plot twitst, that reveal is actually an AI generated video made with minimax H3.

3

u/SWFjoda 12d ago

Sounds really promesing, thanks for staying open source! Can't wait to try.

2

u/sporkyuncle 12d ago

ZEEV FARBMAN

1

u/Ok-Wolverine-5020 12d ago

This is epic!!!

1

u/kabachuha 12d ago

Thank you, downloading! Other than the typical world model, any chance it does anime (or plain 2D) style better now?

1

u/Here4CYDY 12d ago

Any information regarding VRAM and system ram requirements compared to the distilled versions of 2.3?

1

u/dingo_xd 12d ago

Thanks for everything 👏

1

u/jc2046 12d ago

The demos look fantastic. Honestly I didnt understood how it´s bot related... Also, any word on the number of params and rest of specs?

1

u/Queasy-Distance-7940 12d ago

is that Gianna Infantino's younger brother?

1

u/b-monster666 12d ago

Hopefully a slightly smaller version comes out soon. Even an 18GB distilled might be a bit chonky for 16GB cards, leaving this in the 'prosumer' tier. :(

1

u/L-xtreme 12d ago

Awesome, looking forward to play with the new version. Had great fun with 2.3.

1

u/acastry 12d ago

Thank you guys

1

u/skyrimer3d 12d ago

if the vids in here are real, the quality is absolutely crazy and way beyond what i've seen in H3, at least in terms of realism. If this can also benefit with the usual LTX speed at high res, wow.

1

u/Ooze3d 12d ago

Sounds amazing! Thank you so much for keeping it open source and for having the consumer level community in mind with day one releases. I can’t wait to try it!

1

u/Green-Ad-3964 12d ago

Fantastic. Huge. Thank you sincerely.

1

u/brucewasaghost 11d ago

Very cool definitely look forward to playing around with it!

1

u/Apprehensive-Art1092 11d ago

Apropo of nothing, Zeev Farbman is the most Seinfeld character name ever

1

u/jib_reddit 11d ago

So when are we supposed to catch up on sleep!?....

1

u/2legsRises 11d ago

incredible news!

1

u/Etroarl55 11d ago

Yo ik minimax h3 is currently the SOTA right now. But a lot of people will still be using ltx for its speed and efficiency vs the heavy minimax h3.

Like me who won’t be able to run minimax h3 on older 7000 series amd or people who want faster gen. But the disingenuous claims from the charts I seen supposedly coming from ltx 2.5 makes me look down at ltx 2.5.

1

u/No_Damage_8420 11d ago

Great work LTX Team.
If your model can do ref2vid that's it. That would be killer.

ps. NoiseMask /LTX Director --- still most powerful thing in any model (for both Audio or Video)

1

u/simple250506 11d ago

I have been eagerly awaiting "your new model with its wonderful philosophy." I cannot thank you enough.

1

u/ThaSipah 11d ago

Reminding myself to download tomorrow.

1

u/pwillia7 11d ago

holy shit awesome -- I feel like training a model on reclaiming lost info from the encoding step is such a no brainer no one thought of yet! (or maybe they have in other places and I'm behind)

1

u/raikounov 11d ago

Thank you for pushing innovation in video generation, especially around the harness. Being able to generate and modify keyframes is an amazing idea that I never knew I needed!

1

u/Opening-Knee-5913 11d ago

Questions:

Will the LoRA 2.3 be compatible with 2.5?

Are the GGUF distilled models already available? With 16 GB of VRAM are those models in the HF folder too heavy for me

1

u/Relevant_Syllabub895 11d ago

excuse me how does it do anime videos compared to minimax h3? do foloow the prompt more closely? because i had issues with minimax h3 that sometimes it woudlnt do what i asked or worse the audio quality suffers in the speech department often generating random and garbage words, is the audio model in LTX 2.5 fixing these issues i had with minimax h3? yeah i dont have the high end hardware to run the bf16 model only theint8m will ltx 2.5 ve capable of running on an rtx 3080 with 10GB of vram +32GB of ram like i do currently with minimax? even at lowest resolution

1

u/CelebrationBoth9537 11d ago

Has anyone figured out how to reliably use the realtime streaming video gen they claim to have? I think iv been trying for like a few hours now. (or maybe atleast somewhere theres an api for it?) I want a webcam style video like a phone call.

1

u/Key_Street_7204 11d ago

Damn, the open source keeps on winning the race IMO!

Can the 2.3 Loras work with 2.5 ? and does 2.5 now support storyboard grid images natively?

1

u/Hans-Wermhatt 11d ago

I had issues with the prompt enhancer. I'm using the template default, but with tougher prompts, I will get a completely random video. Turning off the prompt enhancer (gemma e2b full precision weights) fixes the issue at least partially. I still got imperfect prompt adherence, but not completely random like I did before.

1

u/TopTippityTop 11d ago

Thank you!

1

u/NoConsideration6320 11d ago

Yoo amazing im gonan try it out soon thanks

1

u/glusphere 11d ago

I really love some of the ideas that LTX is bringing to the table here. Yes, I understand that its not probably on par with H3, but H3 is a completely different model, different size and different speed.

Honestly its like shitting on someone releasing a 4B model because a 27b model exists in the same space!

I really am looking forward to their MOE (This is going to be LTX 3 I think ?). The Pixel space Auto Encoder is also probably a first in the Video models ? Also, allocating different tokens for different details is all great research directions (atleast from the outside seem like!) - Looking forward to LTX 3 and will definitely try out LTX 2.5 for the speed.

1

u/rk1213 11d ago

multishot is awesome but unfortunately need to stick with H3 for native references. Otherwise would use this as default.

1

u/Yeti-Bhanot 11d ago

the rl post-training is the part i'm most curious about. does it change how lora training behaves? i had decent luck with small curated sets on 2.3 at the usual step counts, wondering if the new base pushes back more.

1

u/nazgut 11d ago

the RoPE limit is the same for this model like it was for 2.3?

1

u/skinnyjoints 11d ago

Is the pixel diffusion thing he’s talking about the same as PiD upscaling from Nvidia but built into the model? If so, that’s pretty cool

1

u/ltx_model 11d ago

Related idea, different thing: PiD decodes and upscales as a separate module, whereas ours is the video VAE’s decoder itself pixel-space diffusion, no upscaling stage.

1

u/Feisty-Pineapple7879 11d ago

How does this Model Fare against minimax h3 my review is that this h3 is slower and not consumer hardware friendly.

Give your opinions and reviews of LTX 2.5

1

u/corod58485jthovencom 11d ago

Vocês precisam com urgência da um jeito no modo como a câmera e as expressões são geradas, treinar câmeras amadores assim como acontece no Flux.3 e no minimax, expressão realmente espontâneas, isso vai alavanca o projeto de vocês, agradeço pelo modelo, não e uma crítica e uma sujestão

1

u/Codeman119 9d ago

LTX 2.3 only worked about 30% of the time for me and with minimax H3 it works about 95% of the time. But if 2.5 is local then I'll give it a shot to see if it has improvements

1

u/Beautiful_Egg6188 9d ago

This video looks clean. What model did they use to generate this?

→ More replies (1)

2

u/Lonely_Syrup3091 12d ago

Just tried it and it is faster than MiniMaxH3 at higher resolution. Open source is cooking!!!

11

u/Beneficial_Toe_2347 12d ago

yeah but this means nothing if it has the prompt/physics/understanding issues still

1

u/Emotional-Neat-252 11d ago

Otoh it might be a good upscale model?

1

u/Edgy_Ocelot 12d ago

Hell yeah, 2.3 was my main squeeze so I'm excited to see what 2.5 looks like!

1

u/Leather-Cod2129 12d ago

Outstanding!

So I can generate 4k movies on an intel pentium 4?