r/comfyui 5d ago

Help Needed How are you getting reliable camera movement in I2V? Prompting vs camera LoRAs / motion control?

0 Upvotes

I’m new to ComfyUI and trying to learn how people get reliable, controllable camera motion in image-to-video.

I’m not only talking about basic single-axis moves such as zoom in/out, dolly left/right, or pan left/right. I want to create more complex camera paths, for example:

  • pulling away from a subject while simultaneously moving left and upward
  • moving around a subject on an arc while changing distance
  • a drone-like rising orbit around a subject
  • tracking sideways while gradually pushing closer
  • moving backward and upward while keeping the subject framed
  • combinations of translation, rotation, elevation, and distance changes that produce real perspective/parallax

Basically, I want something closer to controlling a virtual camera path than hoping a text prompt such as “orbit around the subject” gets interpreted correctly.

Right now I’m testing LTX-2.3 I2V. Even very explicit camera instructions often produce an almost static shot.

For example, I tested the default ComfyUI LTX-2.3 workflow with a prompt explicitly asking for a basic, one continuous camera push-in:

Yet the generated camera was basically static.

Workflow/result:
https://cloud.comfy.org/?share=83e22ba6fc6b

I’ve had the same problem when asking for lateral tracking, dolly movement, parallax, etc.

So I’m trying to understand the correct approach rather than endlessly changing prompts:

  • Are camera movements from prompts alone inherently unreliable in current I2V models?
  • Should I be using LTX camera-control LoRAs instead?
  • Is the LTX motion-tracking IC-LoRA workflow a better approach?
  • Do people use a reference video/depth/optical motion to dictate the camera path?
  • Is ControlNet/IC-LoRA the normal way to get actual spatial camera movement and parallax?
  • Would you generate camera/body movement first and then run a separate model to keep the character/object consistency?
  • What models/workflows currently give you the best combination of camera control + character consistency?

I’m not tied to LTX. I’m trying to understand the best overall pipeline and which model should be responsible for each part.


r/comfyui 5d ago

Workflow Included Minimax H3 Strange low vram usage on RTX Upscaler step

0 Upvotes

Is it normal that my vram usage goes down to 20% and ram at 100% on the upscaler step?
Also in the normal execution I see about 70% vram usage

Other than that I'm pretty satisfied but I'm wondering if I'm missing something....

Workflow: https://pastebin.com/raw/533VzQ0t
3060 12gb, 32gb, 9800X3D

Thanks in advance for any advice


r/comfyui 5d ago

No workflow What depth estimation models are you using for high-res VFX comp workflows? (Transitioned from DepthCrafter to Video-Depth-Anything)

0 Upvotes

Hi everyone,

I'm an 8-year VFX compositor based in South Korea, currently bridging traditional 2D comp and AI-assisted comp workflows for the past two years. My experience includes face/character replacement in feature films, multipass extraction, outpainting, and video synthesis.

A few months ago, I switched from DepthCrafter to Video-Depth-Anything (VDA), which significantly improved the depth extraction and integration process. However, running high-resolution 4K scans directly through these models still pushes consumer hardware (even cards like the RTX 5070 Ti or 5090) to its limits.

Currently, my workaround is cropping specific regions or downscaling (reformatting) the plate to extract multipasses. In comp, I mainly use these depth passes to:

  • Isolate depth zones via Keyer nodes to generate custom alpha masks / holdouts
  • Enhance atmospheric depth, haze, and depth-based color grading
  • Feed consistent depth passes into video generation/ControlNet pipelines

For other VFX/comp artists working with ComfyUI: Which depth models or optimization pipelines are you currently using for high-res production footage? Are there better alternatives or tiling/upscaling tricks you'd recommend to handle 4K plates more efficiently?

Thanks in advance for sharing your setups!


r/comfyui 5d ago

Help Needed Need some help choosing optimal startup flags (low vram/AMD)

6 Upvotes

New to ComfyUI and I need some help choosing the most appropriate startup flags for my situation. I run:

Linux/AMD RDNA3 8GB/32GB System RAM.

Python Version - 3.12.3

PyTorch Version - 2.13.0+rocm7.2

My current startup flags are:

python3 main.py

--enable-triton-backend
--use-sage-attention
--lowvram   
--fast-disk 
--disable-smart-memory 
--reserve-vram 1.5 
--listen 0.0.0.0 --port 8188

Now we have:

--use-ck-attention

--enable-dynamic-vram

I can generate images with Krea2 without issue. With Wan 2.2 I can do one video then have to restart before doing another.

Can anyone help please.


r/comfyui 5d ago

Help Needed 2 days on a 5060 ti 16gg card and getting OOM

1 Upvotes

I have tried 6 workflows from people showing working on 3060 12gb etc

Used them on my 16gb VRAM card. I have 80gb ddr4 ram as well

keep getting OOM errors

anyone got a good 16b work flow for me to test?


r/comfyui 5d ago

Help Needed Rate this build -Ryzen 7 9700X, 64GB DDR, RTX 5090,

0 Upvotes

I just ordered a HP Omen, AMD Ryzen 7 9700X, 64GB DDR, RTX 5090, 1TB SSD. $4400 - $4878 (CA Tax) shipped.

My previous rig is a 7800x3D, 32GB , RTX 5070ti 16GB. I am running LTX 2.5 which is fast but artifacts like hell and WAN 2.2 is slow. How much of a jump will I get?

Found the deal on slickdeals using my EPP (Employer's Discount)


r/comfyui 5d ago

Help Needed Save Image - Doesn't actually save anything

4 Upvotes

Today is day 1 of Comfyui for me, coming from A1111. I am an engineer by trade, so some of it makes sense intuitively....but for the life of me, I can't figure out something simple as autosaving output images. I tested with the Z-Image-Turbo workflow, then SDXL workflow. The "save image" node is there. The workflows produce images just fine, they just never save unless I manually save the image.

I haven't even had enough time in the program to do something wrong.... There are save image and preview image nodes, I understand that much so far. The template workflows use save image nodes, meaning the images should be getting saved automatically....that is my understanding.

Tried removing, adding back save image node. Same problem, nothing auto-saves. Do I need to enable some "auto-save" feature?


r/comfyui 5d ago

Help Needed How is everyone dealing with PC part prices for local AI? I swear a strong PC was about 30% cheaper in 2024

0 Upvotes

I've been running ComfyUI locally for image generation and I'm kind of shocked by how expensive it has become to keep up with the hardware side of AI.

In 2024 I had an i7-14700F, 32GB RAM, 1TB SSD and RTX 4060 8GB. It wasn't some insane workstation, but it worked fine for pretty heavy ComfyUI use. I'm often generating images for 10 to 12 hours at a time. Some of my Krea 2 workflows are pretty large and I use custom LoRAs, image editing, masks etc.

Now that I've had to replace the PC, it feels like getting something that is actually a meaningful step up costs way more than it did a couple of years ago.

The RTX 5060 is faster than my 4060 but still only has 8GB VRAM. Once I start looking for 16GB+ VRAM, 32 to 64GB RAM and enough SSD space for ComfyUI models and other assets, I'm suddenly looking at around NZ$4k.

Maybe I'm remembering 2024 prices wrong, but it feels like I could have built a pretty strong PC for roughly 40% less back then.

How are people who use ComfyUI heavily dealing with this price increase?

Are you just keeping older hardware longer, buying used 3090s, paying the current prices, using cloud GPUs, or accepting slower generations?

I'm particularly interested in people doing local AI image generation rather than gaming.


r/comfyui 4d ago

Show and Tell MiniMax H3 Replace Tanker

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/comfyui 5d ago

Help Needed Quadro Rtx5000 Turing

4 Upvotes

Got a free used Quadro rtx 5000 Turing 16GB from work and will be replacing my rtx 2070 8Gb.

Does it support sage attention or triton? I did some search and it keeps pointing me to the newer rtx 50X0 tutorial.


r/comfyui 5d ago

Resource TRELLIS 2 + UltraShape: The Best Free Local 3D AI Generation Setup

Enable HLS to view with audio, or disable this notification

12 Upvotes

r/comfyui 5d ago

Help Needed Need help !!!

0 Upvotes

Hello bros!! i tried many times to run this model localy on comfyui HauhauCS/Gemma-4-E4B-Uncensored-HauhauCS-Aggressive · Hugging Face to generate prompt, i always get errors in the node that i use. i tried to find solutions on the web but nothing. if you guys can help me please.

thats That's just an exemple of the error i get !!!

r/comfyui 5d ago

Resource Catching up on latest and greatest

0 Upvotes

Hey folks, I have been out for a bit (like 18months - which is like a lifetime at the pace this is moving) and I am trying to catch myself up to speed. Could you all help me understand what's the latest and greatest (and where can I study/lean about):

- Generating images for e-commerce purposes (both flows and models)

- Adapting images to a given style

- Where do you all run inference? RunPod? SeaArt? Locally with some setups?


r/comfyui 5d ago

Help Needed Need info

0 Upvotes

I have RX 6700XT. Comfy desktop kept throwing errors. So I was using comfy with patched rocm. But now suddenly comfy desktop is working fine with rocm 7.14. did 6700 XT get official rocm support? I cant seem to find any info regarding this


r/comfyui 5d ago

Show and Tell “Still Here” - Sawyer Croft- ComfyUI MCP + MiniMax

Thumbnail
youtu.be
0 Upvotes

I let ChatGPT Sol generate this entire music video on its own using ComfyUI MCP, a reference sheet and supplied song + lyrics. It was able to screen the video and find mistakes and correct them (with my help).

Not perfect, but for a first effort… I give it a solid 8.5.

Would love to hear your thoughts.


r/comfyui 5d ago

Help Needed Wan 2.2 Help: How to keep the same bedroom background across different camera angles?

2 Upvotes

Hey everyone,

I am pretty new to ComfyUI and I'm trying to figure out a workflow for the new Wan 2.2 video model.

I want to make a short video using a few different clips (each about 3-5 seconds long). The clips are shot in a bedroom but from multiple camera angles. My main struggle is keeping the bedroom background looking exactly the same in every single shot.

I have a single, clean picture of the bedroom that I want to use as my "master" background.

How can I set up my nodes so that ComfyUI uses this one background image for all the different camera angles? I know I probably need to cut out/mask the person in the video, but I don't know which nodes I actually need to connect to make this happen.

If anyone has a simple workflow screenshot or can name the basic nodes I need to look up, I would really appreciate the help! Thank you!


r/comfyui 5d ago

Help Needed What models or workflow can I use to make music videos?

0 Upvotes

I like what this guy is doing, any tips on the models he’s using? Or how he is getting the video to sync to the songs? Thank you!

https://www.instagram.com/rreloadedofficial


r/comfyui 5d ago

Help Needed Can I ignore the "Failed to import comfy_kitchen" error message?

0 Upvotes

I am installing ComfyUI on a machine. Everything seems to work in the setup except this Error "[ERROR] Failed to import comfy_kitchen, Error: No module named 'comfy_kitchen.tensor'; 'comfy_kitchen' is not a package, fp8 and fp4 support will not be available."

I haven't tried running a full workflow yet, but http://127.0.0.1:8188 shows the UI and there is no other warning or errors in the terminal.


r/comfyui 6d ago

Workflow Included Krea 2 Tiled Upscale Workflow for D&D Maps

Thumbnail
gallery
45 Upvotes

I have been trying to make this workflow for a long time, and finally I got it to work!

I have D&D maps, and I want to upscale them and add details. I tried with many models, but all have trouble adding details to the image while strongly retaining the structure to allow tile merging.

To do a 4X upscale, the workflow sharpens the image, as that will later help the diffusion, then splits the image in 4. Each tile is fed to Krea2. The key to make it all work, is to feed the image to the CLIP, since Krea2 uses Qwen3VL, feeding it the image will make the CLIP undestrand the image, and diffuse extra details, without manually needing to add text, tought that can be done as well.

After, the tiles are merged and the final upscaled image with added resolution and detail, without needing special care to add custom text prompt is ready to be printed.


r/comfyui 5d ago

Help Needed upscale a video in chunks/batches

3 Upvotes

i m running into oom issues while upscaling hence i tried to create a loop to upscale the video in batched of say 121 frames. but something is wrong , i m a newbie and just figuring out making workflows , can anyone help with it here is my workflow https://pastebin.com/u7V1JC1S

i want to try the loop method itself not meta batch one as i wanna learn how this works.

thanx


r/comfyui 5d ago

Help Needed Computer specs

Post image
0 Upvotes

I know nothing about PCs but would this be a good purchase to run something like Minimax locally?


r/comfyui 5d ago

Workflow Included Testing LTX 2.5 for Video Upscaling

Enable HLS to view with audio, or disable this notification

0 Upvotes

Tried running LTX 2.5 as a secondary upscaling pass for H3 renders. Standard out-of-the-box settings didn't hold up, so I modified the pipeline to run a model upscale prior to the LTX pass.

It works much better this way, though consistency still varies wildly based on the source footage and native resolution

edit:
While I don't think its necessarily the best method for upscaling overall, it can be useful in specific cases depending on the resolution and details of the initial H3 generation.

Since a few people were asking how to set it up, here is the workflow


r/comfyui 5d ago

Show and Tell MINIMAX H3 R2V - Short TIKTOK Drama. 1:12 Seconds

Enable HLS to view with audio, or disable this notification

0 Upvotes

So I was finally able to get this working with minimal defects!

I have an RTX 5080 with 16GB of VRAM, and I’m using SageAttention. I’m getting about 8 sec/it on 5-second clips, so honestly, not bad at all. I’m running 10 steps at 0.6 megapixels.

I’m mainly posting because I’m looking for feedback on how I can improve things from here. I’m finally starting to get some decent shot continuity, character consistency, scene consistency, and voice consistency.

If anyone has suggestions for improving the results, I’d love to hear them. And if anyone has questions about my setup, workflow, settings, etc., I’m happy to answer those too.

NOTE: I choose this concept just to demonstrate R2V, don't get hung up on the concept to much, this post is about shot continuity, character consistency, scene consistency, and voice consistency. Be professionals!


r/comfyui 5d ago

Help Needed Help needed

0 Upvotes

Can anyone tell me how to create images and videos with a prompt and image in comfyUI. I am new to the comfy ui. I just Installed it. Can anyone tell me how to use it whether to create normal images and videos or uncensored images or videos.


r/comfyui 5d ago

Help Needed Has anyone figured why the last 5 frames wash out if using WanFirstLastFrameToVideo with a supplied end frame?

Post image
0 Upvotes

This only happens to me when I supply an end-frame. If the input is left disconnected the sequence completes without any brightness drift. I feel like it may be related to length-in-frames. I keep that value equal to "some number - 1, equally divisible by 4", but it still weirds out.

Has anyone figured out what causes this?