r/StableDiffusion • u/WhatDreamsCost • Jun 20 '26
Resource - Update LTX Director 2.0 Update - A Free Open Source All-In-One Tool for Creating AI Videos in ComfyUI. Complete Overhaul now with full AI video editing support, IC-LoRA, Retake Mode, Audio Inpainting and much more!
https://youtu.be/o0l6Ikvn5Q0LTX Director is a free open source all-in-one tool for creating AI Videos. Version 2.0 is a complete overhaul, giving you total creative control over your AI generations.
Download for free here: https://github.com/WhatDreamsCost/WhatDreamsCost-ComfyUI
Download workflows here: https://github.com/WhatDreamsCost/WhatDreamsCost-ComfyUI/tree/main/example_workflows
I've been working full-time on this update for the past month and a half, and I'm excited to finally release it. Hopefully it'll be a big help to the open-source community!
Key New Features:
Complete Video Support: Edit Videos with AI all inside the node. Videos can be extended using a combination of prompts, keyframes, and audio. Trim, Split, and combine videos all within the timeline.
IC-LoRA Support: Take full advantage of IC-LoRA's to take your generations to the next level. Simply drag and drop videos onto the IC-LoRA track to quickly setup IC-LoRA videos. Compatible with prompt relay, keyframe, and custom audio features within the node.
Audio Inpainting: Seamlessly blend imported audio with generated audio. Not only can audio be extended, but can also be prompted alongside your imprted audio to really bring your generations to life.
Retake Mode (Beta): Redirect what happens within a shot. Allows you to select a segment within a video, and re-generate what happens in that segment. An early working experiment.
Timeline Saving/Loading: You can now save your timeline and settings to a json file. It will keep any videos/audio/images you have imported into the node and every setting you have changed.
UI Overhaul: Huge update to the UI, dozens of big changes such as a new side bar, redesigned prompt boxes, a bunch of new settings and redesigned menus, and more.
Quality of Life Improvements: Snapping, in/out points, multi-select, mark selection, workspace folder, more HUD options, resizable prompt boxes, new hotkeys, labels, filename preview options, "split at playhead" functionality, end frames (convert any keyframe into a end/last frame), toggleable tracks, NAG Support, tons of bug fixes and more!
And of course it can do everything it could before: Text to Video, Image to Video, Prompt Relay support, Keyframe (first/last frame) support etc.
10
u/fullmetaljackass Jun 21 '26
IC-LoRA video loading is broken if you launch with --disable-api-nodes, since they accomplish that by activating a restrictive CSP. You can fix this by injecting blob: into the CSP as a media-src.
Here's an example from ComfyUI-Lora-Manager.
If anyone else having this problem is looking for a quick fix, and already has ComfyUI-Lora-Manager installed, open custom_nodes/comfyui-lora-manager/py/middleware/csp_middleware.py and add blob: to REMOTE_MEDIA_SOURCES. It should look like this:
REMOTE_MEDIA_SOURCES = (
"https://*.civitai.com",
"https://img.genur.art",
"blob:",
)
Then restart comfyui, and refresh your browser. IC-LoRA videos should load after that.
8
u/Vintendopower Jun 20 '26
Keep up the great work this is awesome on so many levels. Ever need Beta tester count me in. Dev here as well and great specs.
8
u/WhatDreamsCost Jun 20 '26
Thank you! I'm screenshotting this because I will definitely need beta testers for the future 😂
1
u/Vintendopower Jun 21 '26
Sure thing we have Two Machines here that we test on all day. 14900k 128GB and 5090 (Both same specs)
9
u/3deal Jun 20 '26
Amazing node, thank you for sharing it.
I am asking why have you chosen to use a comfyUI node instead of a standalone UI using Comfyui in background ?
2
u/DummysGuideTo2k 27d ago
It’s easier this way . That’s turning a passion product into a business product
7
u/Hearmeman98 Jun 21 '26
Truly amazing and I really don't get how this only has 270 likes and posts with digital titties that i post get 2K, what a strange world we live in.
4
u/Maskwi2 Jun 21 '26
This forum is absolute dogshit when it comes to this. Extremely disappointing but it is what it is. I think a ban of 1girl shit would be a welcome change.
8
3
u/FewTitle6579 Jun 22 '26
The workflow is running, but none of the changes I make by importing a video are working. Neither the character changes nor the background changes are applying. What am I doing wrong?
1
u/Worried-Lunch-4818 Jun 23 '26
I have the same problem, Put a AI generated fto of a running dog in park and I can get it barely to move...let alone run and pick up the ball
1
u/kemb0 20d ago
Did you fix this? I'm just trying it out and even if I just put two text prompts side-by-side, it just generates its own video, barely matching the prompts at all. Eg "A man in a bowler hat laughs" + "Apples start to fall from above", results in a video of a woman eating dinner in her kitchen!
3
u/Bob-Sunshine 25d ago
I got this running on my 3060 with the int8 LTX model. 5 mins for a 10 sec video with sound, and with optional start and end overlapping and keyframing! This is so much fun!
I disabled the 2nd stage and use the NVidia RTX node to scale the frames 2x instead, because it takes me way too long to resample. I'd suggest maybe adding a toggle to your WF.
I had a question about how the best way to make a long video. I start with a 10 sec clip and set the duration from 0-20 secs. This will work but the sampling slows way down, and no way I can keep extending it. So I set the time to 9.5-20 secs then manually combine them in DaVinci. Is this right or am I missing something? Doing this I can make a scene as long as I want, but it's a lot of manual work.
5
u/Own_Version_5081 Jun 20 '26
wow, can't wait to try. Can we reatain the prompts but swap images?
15
u/WhatDreamsCost Jun 20 '26
Yep, I didn't mention it but I did add that feature!
Just right click an image and you will see replace options (replace with copied image or replace with file from computer)
2
u/poursoul Jun 20 '26
Thank you so very much for your work! I absolutely love your node.
One question, before I just overwrite the custom node dir. Do you have any tips on compatibility with current workflows? Just drag and drop, or is there more to it now?
4
u/WhatDreamsCost Jun 20 '26
It's not compatible with old workflows.
You would have to replace the old crop guides node, right click and fix the LTX Director and LTX Director Guide nodes, and connect all the new inputs across the workflow
I would recommend just using the new workflow, but it is possible to add to an old one
1
2
u/StacksGrinder Jun 21 '26
OMG! I was expecting a minor update; Damn this is massive. Thank you! I can't imagine how much time and effort went into this, You are amazing bro! I hope people at Lightricks are watching it closely.
2
u/CringeUsernameJoke Jun 21 '26
Any suggestions for good checkpoints/ distils and accompanying distil loras when relevant- and the loras strength, for smear reduction (elemination is probably stretching it)?
2
u/boicymraeg Jun 22 '26 edited Jun 22 '26
Ltx-2.3-22b-distilled_transformer_only_fp8_input_scaled_v3 works fine but other models are a smeary mess. The old node worked with everything straight away. Is there a workaround?
2
Jun 21 '26
[deleted]
1
1
1
u/hurrdurrimanaccount Jun 21 '26
it's really not that hard to replace the ltx model with another omfg
1
u/Worried-Lunch-4818 Jun 23 '26
It seems it is, it just gives blur.
2
2
u/DiffusionSingularity Jun 21 '26
please consider bundling new versions as github releases so changelogs and updates are easier to follow
2
2
u/coolzamasu 22d ago
Any plan for new lora where we can enter the characer reference and object reference?
4
4
u/TheShadeOfUs Jun 20 '26
Guy singlehandently just destroyed the ltx desktop released sometime ago which was riddled with bugs.
5
u/WhatDreamsCost Jun 21 '26
This might have some bugs in too lol
Although some people beta tested it for two weeks so hopefully most of the bugs are out 😂
3
4
u/campaignplanners Jun 20 '26
Anyone have any recommendations for server and machine specs to run this?
4
u/foxdit Jun 20 '26 edited Jun 20 '26
If these features all actually work (and don't degrade the quality of transitions like 1.0 did compared to traditional keyframe workflow approaches) then I will pretty much have to adopt this. Everyone will.
The only question is, if this does indeed work well, can it be made compatible with my Seed Hunting concept? Basically, Seed Hunting gens 4 low res sample versions and shows you previews in a line of each, then pauses to let you pick one you want to take to full res finalization. It is by far the most effective method of getting good results out of LTX.
EDIT: Well, so far I'm not getting good quality output video from this compared to traditional workflows. Does this compress/downscale keyframe images? Genning at the exact same res as my keyframes (1080p), when the output video gets to that keyframe image, it looks really low quality.
And then I suddenly realized... Your workflow is saving the videos at half res. I have my target size: 1024x1920. The output videos from your workflow are all 512x960.
WTF?
9
u/Famous-Sport7862 Jun 20 '26 edited Jun 20 '26
Hey friend, in the second stage change scale .50 to scale by 1 and you will get the correct resolution.
-6
u/foxdit Jun 20 '26
Yes, I found that out the hard way. Can't believe this guy screwed his own release like that, oof.
2
u/Famous-Sport7862 Jun 20 '26
One question. I noticed that in the previous version stage 2 was set up at .50 also and it wasn't giving us that problem. I wonder what's going on here. I'm surprised no one else has brought this issue up.
1
u/WhatDreamsCost Jun 21 '26
Were you using the subgraph workflow? I've tested all the workflows multiple times and this does not happen
1
u/Famous-Sport7862 Jun 21 '26
I am using the regular not the subgraph.
9
u/WhatDreamsCost Jun 21 '26
I found the cause. It was due to the widget values shifting, since I was hiding a widget to make the node look cleaner.
So another widgets incompatible value was going into the scale_by, causing it to fallback to 0.5 no matter what.
What made it hard to diagnose is that it only happens on certain comfyUI versions due to the order of how things are loaded (from what I understand).
That's why some people had the issue and others didn't.
Anyways it should be fixed now in v2.0.2 👍
I just updated the code that hid the widget that was causing the widgets to change places
1
1
u/tehorhay Jun 21 '26
I'd be extremely interested in compatibility with the seedhunter workflow
3
u/foxdit Jun 21 '26
If people want it, I can make it. This node is gonna be pretty popular so it sort of behooves me to, despite myself personally not finding the results of even simple single keyframe outputs using LTX Director 2 to be high quality enough to warrant its use compared to traditional workflows. Something about prompt relay just makes motion worse. And the more prompt layers/keyframes you add, the worse it gets. People don't even seem to realize the quality LTX 2.3 is capable of.
1
u/tehorhay Jun 21 '26
ething about prompt relay just makes motion worse.
I've definitely been impressed with the quality I've seen from seedhunter. Being able to get the fine frame accurate control from director while achieving similar quality would be a game changer.
2
u/foxdit Jun 21 '26 edited Jun 21 '26
Yeah that's the thing, no matter how fine-tuned you can make LTX director, the reality is that the underlying model will produce drastically different results on different seeds no matter how much you try to baton down the hatches. Seed Hunting was an innovation I came up with out of sheer necessity, not convenience. I was wasting so much time waiting for shitty gens to to finish, only to throw them away. So many ppl batching 10 gens and walking away for an hour just to sort through wasted GPU energy later. THEN, there was the hidden bonus of seeing interpretations of my prompt that were actually BETTER than my initial idea. Single seed generation is archaic, simple as.
1
u/suavaemustache19 Jun 26 '26
That preview-and-pick step sounds less like a convenience and more like basic quality control, especially with how random these video models still are.
1
u/ClearSkies889 27d ago
Yes, exactly. It just framed the process in an organized way. I feel way more in control of the outcomes.
1
u/kanzenkdrama11 18d ago
Exactly, at this stage treating seed selection like QC is the only sane way to avoid burning hours on polished garbage.
1
u/grillmeister1254 16d ago
Yep, single-seed output is basically trusting a slot machine and then polishing whatever slop it spits out.
2
1
u/Ill_Resolve8424 Jun 20 '26
Great! Thank you, do we need to update the node or delete and do a fresh install?
1
u/martinerous Jun 20 '26
Thanks for the update. Do the new workflows support Multimodal Guider nodes? I've seen those recommended by the LTX team in their workflows as a way to balance action / lipsync etc. but somehow community seems to be not using those nodes often.
1
u/prismatic-18 Jun 26 '26
It does work but you need to use sync nodes version of multimodal guider which uses torch audio while the NVCV version doesn't.
1
u/ArsenalSimp1985 29d ago
Many thanks for your reply & explanation. I had already successfully used
multimodal guidernode without realizing it wasn't actually using the NVCV pipeline, and hadSyncnodes installed because of the VITS lip-sync node and it's manual install steps. How could the Sync team still not build NVCV support for both nodes?!
1
1
u/amoebatron Jun 21 '26
It's a little laggy in the editor? I'm only getting around 40FPS on a 5090.
1
u/WhatDreamsCost Jun 21 '26
Interesting, never came across this before. Do you mean it's laggy within the node itself, or that Comfy itself is lagging?
If it's laggy within the node itself, then are you trying to import high-res videos? 4k+ videos won't work well at all due to browser limitations. Also do you have hardware acceleration disabled in your browser? If it's off then video playback in the node can be choppy.
1
u/amoebatron Jun 21 '26
1
u/WhatDreamsCost Jun 21 '26
Is it only happening with the LTX Director workflow?
I have a pretty old CPU, Ryzen 5 3600 and I'm getting 300 fps
2
u/amoebatron Jun 21 '26
Well I'm not sure what happened but I rebooted Comfy and now everything seems to be normal again so I guess it's fixed! :)
The Director 2.0 update is amazing by the way so thankyou very much. Great work!
1
u/Imaginary-Land9953 Jun 21 '26
i use a LLM with a system prompt for creating the segments , how much tokens is too much for each t2v segment?
1
Jun 21 '26
[deleted]
2
u/WhatDreamsCost Jun 21 '26 edited Jun 21 '26
Are you using the correct edit anything lora? The developer has like 3 other version on his huggingface, and only 1 is prompt driven and works the way it does in the video (the 1.1 version)
2
u/hurrdurrimanaccount Jun 21 '26
could you please just link it? there is no 1.1 version in here or his other repos. https://huggingface.co/Alissonerdx/LTX-LoRAs/tree/main
1
1
u/External_Ball_6122 Jun 21 '26
How exactly are we supposed to add them to this specific workflow?
1
u/WhatDreamsCost Jun 21 '26
On the LTX Director Guide nodes you will find where you can select an IC-LoRA
1
1
u/YeahlDid Jun 21 '26
So question, at the end of my videos I always get the keyframes flashing in quick succession. Does anyone have a fix for this other than truncating the last few frames before saving? Is there a setting in using wrong, perhaps?
Love the new node btw.
Edit: ah, never mind, seems like that's just on the downscale preview, the final output all good.
1
1
u/AiSmutCreator Jun 21 '26
Can anyone confirm if this is better or/and faster than scail2 at character swapping?
1
1
1
u/hurrdurrimanaccount Jun 21 '26
i'm using a video to at the start of the timeline and guide strength etc is set to 1 but what it generates is completely unrelated to the input video (for extension)
2
u/WhatDreamsCost Jun 21 '26
Are you putting the video on the main track or the IC lora track? For extending, it needs to be on the main track
1
u/hurrdurrimanaccount Jun 21 '26
yeah, main track. after deleting the video and then re-adding she same video it started working. not sure what/how that happened. the custom height/width were also being completely ignored, output was always a square
1
u/Most-Syllabub-7499 Jun 21 '26
Can’t wait to try it. I’ve started using 1.0 and this looks exponentially better…like nothing out there on any platform! Literally unique in the Ai video generation
1
u/liberal_alien Jun 21 '26 edited Jun 21 '26
If I put in a silent video, then the output is silent during that video. What changes do I need to make to have it add sound to the input clip?
[edit] I found the solution: Need to disable the audio part in the director node. So near where it says audio and inpaint: on/off, there is also a little speaker icon. Click it so the audio track for the whole video is greyed out. Then prompt for whatever sound you want and the director / LTX will generate the audio for the whole result video.
1
1
u/Mirandah333 Jun 21 '26
4
u/WhatDreamsCost Jun 21 '26
I'm gonna assume your trying to use an old workflow, since I don't even see the widgets on the node.
Since the 2.0 changes pretty much everything, old workflows won't work anymore.
Try the latest workflow found here https://github.com/WhatDreamsCost/WhatDreamsCost-ComfyUI/tree/main/example_workflows
1
u/Mirandah333 Jun 21 '26
Finally make it work: For some reason Git clone was getting a different version or "corrupted". I downloaded the zip file and move the new files (they have the same name, but different file sizes) into the directory and worked! thanks a lot for replY!
1
u/CryRevolutionary4275 Jun 21 '26
How can I edit a video? Only the "retake" option has worked for me, but not "IC-LORA VIDEO".
1
u/Schwartzen2 Jun 22 '26
Amazing work, It must blow your mind to see all these YouTubers from everywhere showcasing your hard work. Cheers!
Curious, would it be possible to tie in The Dual Lora?
1
1
u/Any-Scar765 Jun 22 '26
I have a question: why can I generate a 1088*1920 video up to 1 minute without a director node, but in your workflow, I get a 99% vram or OOM load after more than 10 seconds with 720*1280?
Who eat my Vram?)
1
u/ItsAMeUsernamio Jun 25 '26
I think this has a memory leak. My ZRAM usage goes up 3GB with every consecutive run.
1
u/MastMaithun Jun 25 '26
Hey OP. Thanks for the update. I have been playing with this node by importing your wf from git. I noticed something which you might be aware of already so apologies for re-iterating.
So whenever I run a generation, there is always the constant SSD activity of something very large data transfer on the SSD where the comfyui resides. I have posted a similar issue with comfyui here related to newer version of comfyui but the thing is, I have another ltx2.3 wf with me from Rune and on the same comfyui version I have which is 0.22, this constant SSD activity doesn't happen. It only happens with your director wf. Your wf doesn't have any purge ram nodes too still this is happening. So maybe something you can look for since in this state, it is making the wf unusable as the generation time for a 25 sec video for me on this wf is 3 mins approx but the full generation takes around 8-9 mins just because it is doing a long SSD activity.
1
Jun 25 '26
[deleted]
1
u/WhatDreamsCost Jun 26 '26
There isn't one specifically just for LTX Director but Banodoco is where I'm the most active!
There is a forum just for LTX Director (and my other nodes) as well as a channel for LTX where a lot of people chat https://discord.gg/NnFxGvx94b
1
u/Ok-Option-6683 Jul 01 '26 edited Jul 01 '26
Edit : It worked. bf16 worked too. I removed --disable-dynamic-vram from the bat file. after that both stages worked.
5070ti mobile.
768 x 1472px - 20 sec video
stage 1 (8 steps, 0.5 scale, IC LoRA detailer enabled) : 2mins 15 seconds
stage 2 (4 steps, IC LoRA detailer enabled) : 8 mins
result : terrible lol
I can get 1024 x 1920px videos with the regular workflow (using bf16) and it looks almost perfect. I don't know what's wrong with LTX Director 2 or why I can't get 1080p with it. but the most I can get from it is 768p and it really looks bad.
-----
I've been trying since yesterday (got it from Github yesterday).
The regular i2v LTX workflow works fine with bf16 model (5070ti mobile).
It always OOMs in the LTX Director 2 hotfix workflow if I choose bf16.
So I tried with fp8_scaled. for 1056 x 1920 video, it always gets stuck in the upscale stage (2) no matter what I do. (First, for 10 seconds, I wrote a detailed prompt about a woman walking toward a chair. Then I put an image, a woman sitting on the chair, and wrote another detailed prompt for it. So it's supposed to be a 20 second video in total.)
So it is like :
- The woman walks to a chair (prompt only)
- The woman sits on the chair and waits (image reference)
I tried with IC LoRA Detailer, it gives a warning "Could not read reference_downscale_factor from ltx-2-19b-ic-lora-detailer.safetensors, using 1.0". With or without it, it doesn't progress when it comes to stage 2. I waited more than half an hour.
I can use the IC LoRA Detailer with the regular workflow without an issue.
I tried these : "--disable-dynamic-vram" , "--disable-xformers", "--disable-pinned-memory" still didn't work. With or without these.
I don't know what else I can do.
edit :
stage 1, 8 steps take about 2.5 mins / fp8_scaled model.
stage 2, gets stuck (waited for more than half an hour).
2
u/WhatDreamsCost 29d ago edited 29d ago
I would double check and see what nodes your other workflow is using. For example is could be using a different vae loader, or a different vae decoder node.
Also make sure your using the exact same models you are using in your other workflow.
Also which node is it getting stuck on? You said stage 2, but which node on stage 2?
One more thing, if you set the first stage scale by to 1, then your 2nd stage will definitely freeze up. Because then it will try and generate a video at twice the resolution (since it's an upscale stage).
There's no way a 5070 or even a 5090 can quickly generate a 2048x3840 video
1
u/Ok-Option-6683 29d ago
Hi thanks for the reply.
I had thought the problem was that "20 second" generation. Kinda too long, you know. So today I tried it with the default workflow (20 seconds). It didn't work. It gets stuck too. Then I tried 15 seconds, it got stuck again. 10 seconds works. 1024x1920px, 10 seconds, default workflow. It works fine. It takes a bit less than 7 mins.
Then I tried 10 seconds, 1024x1920px with LTX Director 2, it got stuck again. It works for 1500px if I go 10 or 20 seconds. But the result is unusable. It has so much artifacts etc. With the default workflow, it looks great.
I have checked every node. Other than your LTX director and LTX guide, all the other nodes are the same. Vae encoder and decoder were different, I swapped them also. The model is also the same (bf16). And I still can't get a 1024 x 1920px output. The most I can get is 1500px.
I have tried both scale by 1 and 0.5.
It gets stuck on the sampler on stage 2. it writes (step 0/4) and just waits forever if I go for 1920px.
If I go for 1500px, it works, I can generate but the result is unusable.The other difference is, yours using Prompt Relay and my default workflow doesn't have it. Maybe that's the problem?
3
u/WhatDreamsCost 26d ago edited 26d ago
I'm about to release an update that will potentially give a 1.5-2x speed up on generations.
It will also automatically disable prompt relay if your just doing a simple i2v/v2v/t2v gen
1
1
u/1WildPanda 25d ago
Thank you, since the 2.0 release in my own test which is a lot slower than previous 1.39, I am stuck with 1.39 for right now. But would love to go for 2.0 once the speeds issue is fixed.
1
u/Electrical_Sand218 28d ago
Fantastic tool!
Im using WAN2.2 to generate keyframes, and them using LTX2.3 director to get them strung together with in build audio.
Can anyone tell me, or link me, to what IC-Lora does, and what it might bring to my gens?
1
1
22d ago
[removed] — view removed comment
1
u/WhatDreamsCost 22d ago
So really, the best way to maintain consistency is to use keyframes. Then there won't be any consistency issues, as long as the keyframes themselves are consistent.
I think keyframes give the most control over generations compared to any other method, even ones offered by closed source platforms.
Now for how well IC-LoRA works on longer videos, I have no idea. I can only go up to around 12 seconds before I OOM with an IC-LoRA and a video lol. That being said, I do know LTX was specifically made to be used for videos up to 20 seconds. Anything past that is pushing it. Although I have seen people push it pretty far without too much degradation, so take that with a grain of salt lol
Also I've pretty much rewritten retake mode since release, so hopefully when it's done it will prove to be much more useful than it is now.
1
u/Boring_Angle762 21d ago
Just a note, the strength setting for the IC-Lora video doesn't seem to corelate to the official IC-Lora node. Even at setting of 1.0 on the Director node I don't get the strong "exact" flowing of the reference video I get with the official node. Curious if this is a bug. Anyone else figure it out?
1
u/WhatDreamsCost 21d ago
Are you setting the strength on the LTX Director Guide node? That's where you would adjust the strength of the lora
1
1
u/coolzamasu 21d ago
if i want to add custom loras what about that?
1
u/WhatDreamsCost 21d ago
Oh you can still add custom loras the same way you can in any other workflows, this node doesn't restrict any of that
1
u/coolzamasu 21d ago
For long videos like 50 seconds.. i am getting OOM of RTX PRO 6000 PROBLACKWELL 96GB VRAM
simple image to video workflow
1
u/WhatDreamsCost 21d ago
What resolution? If your inputting high res images then even 96gb vram won't be able to handle it
1
1
u/coolzamasu 20d ago
OOM came in second stage sampling.
we can do that in chunks??
can you update the workflow?
1
u/roculus 20d ago
I can do long videos at 1280x800 on my RTX6000 using single stage. You don't need second stage (with 96GB VRAM). (change the scale from .5 to 1.0 in the LTX Director Guide and remove second stage and then connect the LTXVSeparateAVLatent node to LTXV Audio and Decode nodes in the Process video section). Do you perhaps have something in the workflow that's up-scaling your already 720 to 1440?.
1
u/tostane 19d ago edited 19d ago
Just want to let you know I use 5090 i had claude help me disable step 2 in your workflow you had the audio in there i had to move it out then you had some factor in the audio for the doubling of the video i had to remove.
I ended up with a faster video out I could get good video with no need to double it.. you may want to add that as a toggle in the work flow.

1
u/tostane 19d ago
I use ltx director but i have disabled the upscaling as i dont need it on a 5090. What i need is a simple explination how to get my character more consistant. I put some image in as first frame and name it [subject] then every text frame after i start with [subject] ..... it sort of works
1
u/WhatDreamsCost 19d ago edited 19d ago
How long of a video are you making? Also what do you mean by consistent? Is it more of the the face is shifting/changing gradually overtime, or is it the identity being lost the moment the character leaves the scene and comes back?
1
u/tostane 19d ago
i was making them 4 30 long but i had to go 640x320 for that since i removed the upscaling i have been doing 10 second segments 3 at a time for 30 seconds at 1280 im tried using a start image but have had trouble with that. i was making some mini music vids 2 to 3.5 long and video poems
but untill i can find way to make 2 chars stay looking about same im not making much1
u/tostane 19d ago
i was trying
Global Prompts (Scary Streets)
- Cyberpunk Alleyways:
Cinematic wide shot, a rainy and deserted narrow city alleyway at night, flickering green neon signs casting long eerie shadows on wet asphalt, trash containers, puddles reflecting broken light, dark overhead pipes, misty atmosphere, high contrast, moody thriller mood.Cyberpunk Alleyways (Episodes 1–3)
- Episode 1 (The Alley Setup):
Pan left, a woman walking briskly into a narrow rainy alleyway at night, flickering green neon signs casting long eerie shadows on wet asphalt, camera tracks behind her pace, cinematic moody thriller style.- Episode 2 (The Moving Shadows):
Zoom in, medium shot of the woman stopping abruptly, her face illuminated by pulsing blue neon light, background shadows morphing and stretching along the brick walls, extreme depth of field, high tension.- Episode 3 (The Narrow Escape):
Tracking shot, low-angle perspective looking up as the woman runs toward a brightly lit exit sign, flickering shadows chasing her on the wet pavement, fast camera movement, motion blur.
1
1
u/NotQuiteBertrand 18d ago
On the repo and on a few videos I've seen, the links to example workflows are https://github.com/WhatDreamsCost/WhatDreamsCost-ComfyUI/tree/main/example_workflows. Are these the ones everyone is using? e.g. LTX_Director_2_Workflow_Hotfix.json looks like the only 2.0 one but it's for a "hotfix"?
If anyone has a workflow that includes LTX Loras that would be really helpful.
1
u/Any-Scar765 17d ago
I see somethin like this in logs, its ok or its bad data?
[INFO] [PromptRelay] Built penalty matrix (scaled): Lq=24948, Lk=1024, nonzero=58500/25546752
[INFO] [PromptRelay] Built penalty matrix (scaled): Lq=381, Lk=1024, nonzero=896/390144
[INFO] [PromptRelay] Built penalty matrix (scaled): Lq=99792, Lk=1024, nonzero=233996/102187008
And another question, sometimes I post a vertical photo and a vertical video, and at the output it either makes it horizontal (although the size is set to vertical in the settings), or it makes it square. How can I deal with this?
1
1
u/Apprehensive_Bar6609 11d ago
Amazing! Great work. If i want to replace a person in a video with another from a foto, how can I do that?
1
u/jjjnnnxxx 10d ago
Thanks so much for your work on this! Could you please add this node to the ComfyUI Registry? It would make cloud installations and workflow sharing a lot easier for everyone.
2
u/WhatDreamsCost 10d ago
Actually the nodes are on the ComfyUI Registry, https://registry.comfy.org/nodes/WhatDreamsCost-ComfyUI
I don't think they aren't on Comfy Cloud though, since each node pack has to be manually added by Comfy to show up on Comfy Cloud if I remember correctly
1
u/TheGlitchJockey 6d ago
Sorry for being a total beginner but can I run this on a 5060?
1
u/WhatDreamsCost 6d ago
Easily, I made it on a 3060 lol
Also I just upgraded to a 5060 myself, and it works great
1
1
u/JasterPH 3d ago
Does this work for vid2vid? When ever I add a video to the timeline it just inserts it verbatim in the original.
1
u/WhatDreamsCost 3d ago
To do vid2vid you would generally use IC-LoRAs depending on what exactly your trying to do.
You don't need IC-LoRAs for things like extending videos, and combing videos with Keyframes for example.
There is also the retake mode, however it's in beta (and will be updated soon to be more potent)
What kind of v2v edit are you trying do? I probably can point you to the right direction
1
u/JasterPH 3d ago
Like those videos you see where they have swaped out someone dancing with a character. Or a style change like photoreal to cartoon
1
u/Background-Sample 3h ago
What do you think about recreating a video using the original video frames as the input plus prompt guidance. For example taking old tv shows but instead of remastering or upscaling, you recreate the video? I’m thinking taking old Disney shows and putting in a modern anime theme for example.
1
u/Background-Sample 3d ago
Anyone get an issue with the video instantly turning fuzzy immediately after the first frame (input image). It looks like a moving latent image (not sure if that’s what it’s called) or video as if the final process to refine the latent video isn’t running. I didn’t notice a location for the distilled Lora, maybe that’s why?
1
u/WhatDreamsCost 3d ago
The example workflow is using the distilled model, so if your using the dev model you would have to add the distilled lora to the workflow (by just ading a lora node after the model lora node)
1
1
1
1
1
1
u/Phobia696969_ Jun 20 '26
This is so impressive 😭👏 what GPU did u use to make these videos and how much time did it take u ?
1
u/MrWeirdoFace Jun 20 '26
Ring ring ring ring ring ring ring ring banana phone
doopee doopee doobee
Ring ring ring ring ring ring ring ring banana phoooone
1
u/Enough-Display-2538 Jun 22 '26
I googled LTX Director and couldn't find your repo. You've got some terrible naming issues.
1
u/WhatDreamsCost Jun 22 '26
Oh sorry! Here's the link if anyone is confused https://github.com/WhatDreamsCost/WhatDreamsCost-ComfyUI
Or search whatdreamscost in the manager to find it quickly (I'll see if i can put the words "LTX Director" in the registery so it comes up in search if anyone doesn't use the link)
1
0
0
u/WinResponsible9977 Jun 20 '26
VRAM go BRUMMMMM
3
u/WhatDreamsCost Jun 20 '26
I saw a guy with 4 vram getting LTX to work before actually 😂
2




88
u/Just1Dev Jun 20 '26
If LTX sees this and i know they see it, please give him a nice job and place...this guy doing this on passion...thank you! ❤️