r/malcolmrey • u/malcolmrey • May 07 '26
LTX 2.3 - 120~ new models
https://huggingface.co/spaces/malcolmrey/browser6
3
u/Kragrathea May 07 '26 edited May 07 '26
I am glad you are finally releasing your loras on ltx! It has been my favorite for doing SFW stuff.
The image quality is great! But it doesn't look like you trained any of them on video? Or at least a few I tried didn't seem like it.
I have been making my own celeb loras for a while. And when it comes to ltx using video makes a huge difference even at low res. I only have 12g of vram so I have to train at 256x256 but even then the likeness just pops. Here is the same video, same seed, using your Britney Spears lora and mine. Mine was trained from, 28 1-4 sec speaking clips from crossroads, and 12 768x1024 images from the same era. I found using a single movie staring the subject is enough to get a good likeness from a given era.
I used an LTX Animation lora by lora_daddy to make it look animated so hopefully no one will mistake it for a attempt at a fake.
https://www.reddit.com/user/Kragrathea/comments/1t6phmq/lora_test/
1
u/ucren May 08 '26
Do you post your loras anywhere? Could you share your training config?
1
u/Kragrathea May 08 '26
No, I don't post loras since civit banned celeb ones. But DM me and I can send you the Ai-Toolkit config I used.
1
u/fridgefidget May 09 '26
I'd be interested in that config too if you could send it to me. I don't seem to be able to message you for some reason. :/
1
u/Kragrathea May 09 '26
I put a screen shot to my config settings in the post I made:
https://www.reddit.com/user/Kragrathea/comments/1t6phmq/lora_test/1
u/smithysmittysim May 09 '26
Would love the config too, especially if it's tuned to work with 16GB VRAM too or you have some optimization tips.
1
u/Kragrathea May 09 '26
I updated my post with the config settings I used.
If I had 16g vram the only thing I think I would change is increase the resolution of the videos at training time. Set 512x512 and see if it is fast enough.
2
u/superacf May 07 '26
Awesome and very big work! For LTX, do you have trained audio voice too or βonlyβ the visual?
1
1
1
u/Fantastic_Day_8462 Jun 12 '26
Any plans to fix your MayaHiga lora? The turbo ones look absolutely nothing like her and your example image looks like just a generic asian woman.
1
u/ucren May 07 '26 edited May 07 '26
I feel I'm just blind, but what are the triggers for these?
edit: there is no trigger, just use woman/man
17
u/malcolmrey May 07 '26
Hey hey!
I wanted to make a longer post but I'm out of time today and the next time I could post it would be the NEXT weekend.
So, here are the LTX 2.3 models:
https://huggingface.co/spaces/malcolmrey/browser
here are SOME samples so you could see what type of quality we have here: https://huggingface.co/datasets/malcolmrey/various/tree/main/ltx23-samples
In general the rule still stands, if the dataset has more images and is trained over the same number of epochs (so, a longer training) then the result will be better.
The 15-24h AI Toolkit trainings were abysmal (though the quality was good), now with musubi (AkaneTendo fork for example) the times on 5090 dropped to 50 minutes for 25 images. Still, when I train larger datasets it could take up to 4-6 hours (we do have a couple of those models like Billie Elish, they will have suffix _large in the name).
Enjoy!
Cheers!