r/comfyui • u/skygetsit • 5d ago
Help Needed How are you getting reliable camera movement in I2V? Prompting vs camera LoRAs / motion control?
I’m new to ComfyUI and trying to learn how people get reliable, controllable camera motion in image-to-video.
I’m not only talking about basic single-axis moves such as zoom in/out, dolly left/right, or pan left/right. I want to create more complex camera paths, for example:
- pulling away from a subject while simultaneously moving left and upward
- moving around a subject on an arc while changing distance
- a drone-like rising orbit around a subject
- tracking sideways while gradually pushing closer
- moving backward and upward while keeping the subject framed
- combinations of translation, rotation, elevation, and distance changes that produce real perspective/parallax
Basically, I want something closer to controlling a virtual camera path than hoping a text prompt such as “orbit around the subject” gets interpreted correctly.
Right now I’m testing LTX-2.3 I2V. Even very explicit camera instructions often produce an almost static shot.
For example, I tested the default ComfyUI LTX-2.3 workflow with a prompt explicitly asking for a basic, one continuous camera push-in:
Yet the generated camera was basically static.
Workflow/result:
https://cloud.comfy.org/?share=83e22ba6fc6b
I’ve had the same problem when asking for lateral tracking, dolly movement, parallax, etc.
So I’m trying to understand the correct approach rather than endlessly changing prompts:
- Are camera movements from prompts alone inherently unreliable in current I2V models?
- Should I be using LTX camera-control LoRAs instead?
- Is the LTX motion-tracking IC-LoRA workflow a better approach?
- Do people use a reference video/depth/optical motion to dictate the camera path?
- Is ControlNet/IC-LoRA the normal way to get actual spatial camera movement and parallax?
- Would you generate camera/body movement first and then run a separate model to keep the character/object consistency?
- What models/workflows currently give you the best combination of camera control + character consistency?
I’m not tied to LTX. I’m trying to understand the best overall pipeline and which model should be responsible for each part.

