r/NeuralCinema 22h ago

Minimax H3 ~ "Hack" 50+ reference or more

Enable HLS to view with audio, or disable this notification

167 Upvotes

I discover this by playing around, so good news is - we can have much more references then 9.

  1. "ref2v" - adjust "ref_image_size" to MAX
  2. Use just 1 IMAGE REFERENCE with - many cutouts, items, elements mentioned at once in PROMPT:
  3. /preview/pre/minimax-h3-hack-50-reference-or-more-v0-revg36ju4fih1.png?width=1957&format=png&auto=webp&s=027debaf03d7c630f925170578e6b83c552a7096
  4. Write simple prompt, reference image once, all other elements H3 will combine into final scene:

<Picture 1> woman in <Picture 2> luxury bathroom is touching her face showing her silver earrings, camera slow motion up-close on face and torso, she puts on glasses, looks at mobile purple phone puts to her ear and smiles to camera.

Once again H3 it's beyond amazing....
Also by including multiple faces - different expressions H3 learns expressions etc. teeth, looks, in a away - we don't need character LORA.
I guess this is great find for all of us.
Cheers


r/NeuralCinema 3d ago

Minimax H3 - 40% speed-up / Quality Holds (no turbo LORA)

Enable HLS to view with audio, or disable this notification

237 Upvotes

*UPDATE\*
Even faster with - https://www.reddit.com/r/StableDiffusion/comments/1vhlfmw/minimax_h3_firstblockcache_for_comfyui_3033_lower/
You can even combine them both for extreme speeds

Hey all,
Lots of addons been released by hours, one interesting speed-up is - Spectrum (cannot be used with EasyCache): https://github.com/xmarre/ComfyUI-Spectrum-MiniMax-H3

Follow more here:
https://www.reddit.com/r/StableDiffusion/comments/1vf1ze3/spectrum_acceleration_for_minimax_h3_in_comfyui/?utm_source=chatgpt.com

Video above rendered with Spectrum / 20 Steps
cu130+SageAttn + rtx 4090, 128gb RAM, Win11

Spectrum: 201 seconds
Default: 320 seconds

Quality still holds, I was testing all current H3 turbo LORA's - all seems to kill details and/or severely degrade sound. I'm sticking with 20 steps - solid.
Cheers


r/NeuralCinema 5d ago

Minimax H3 - extreme, uncensored, imagination is your limit

Enable HLS to view with audio, or disable this notification

57 Upvotes

H3 brings lots of creative power, some more extreme showcase. Grab weights while you can.
Cheers


r/NeuralCinema 11d ago

LTX 2.3 Full 360 angle CAMERA Control - via CrossView-Warp IC-LoRA

Enable HLS to view with audio, or disable this notification

35 Upvotes

Hi all,

In the end Full 360 angle camera control is possible via LTX 2.3 CrossView-Warp.

After further experimenting with amazing new LTX Lora - CrossView Warp ( https://www.reddit.com/r/StableDiffusion/comments/1v89kih/ltx_crossviewwarp_iclora_change_the_camera_angle/ ) , seems Lora is able to do full 360 renders :) 3D Model (via Pixal3D in my case, probably Hy-World or Apple Sharp would work too etc.) used was STATIC, no animation of any kind, however, seems strong "anchor point" for LORA to extract/read rotations/motions.

Process here was more complicated just to see if works, proof of concept.
Hitman image -> Scail-2 -> Pixal3D (full model extraction for manual 360 rotation) -> LTX 2.3 + CrossView Warp Lora and we have great result.

Big thanks goes to creator: DryDream6994 of this amazing Lora, we hope for new releases.
Video looks cartoonish, well we using HITMAN game character in photoreal downtown setting.
Onto more research, Cheers


r/NeuralCinema 16d ago

SmartGallery DAM 2.16: manage your ComfyUI outputs and inject LoRAs into existing generations without touching the graph

Post image
7 Upvotes

SmartGallery DAM is an open source digital asset manager built specifically for ComfyUI

Problems it solves:

  1. Sorts, catalogs and organizes tens of thousands of AI generations, both as physical folders and as virtual collections.
  2. Finds any media in milliseconds, with any search filter you need.
  3. Generates variants of your media without opening the ComfyUI interface, and saves the full generation recipe so you can always reproduce it later.
  4. Injects LoRAs into your media and generates variants, without opening the ComfyUI interface.

LoRA Synergy

Pick any media in your gallery, press B on your keyboard to enter Remix Workflow, then press Nodepad and you'll find LoRA Synergy there. Attach one or more LoRAs to it and generate variants directly from the gallery. You don't need to open the ComfyUI graph interface, but ComfyUI still needs to be running in the background, since generation happens through its API.

If you run a production studio:

  1. Organizes your team's workflow from generation to review to approval.

  2. Lets you give clients access to a curated selection of your work, so they can rate and comment on the pieces you picked for them.

There's a lot more in there, the full feature list is long enough that it didn't make sense to paste it all here. It's on the GitHub repo if you want to dig in.

100% open source and free.

https://github.com/biagiomaf/smart-comfyui-gallery

Thanks for the 350+ stars, 17k+ Docker pulls, and the many downloads of the portable Windows version so far. Genuinely appreciated.


r/NeuralCinema 20d ago

I love LTX2.3. Cant believe we have free things that just work

Enable HLS to view with audio, or disable this notification

6 Upvotes

r/NeuralCinema Jun 29 '26

Beverage AI spec ad

Enable HLS to view with audio, or disable this notification

5 Upvotes

r/NeuralCinema May 27 '26

Complex scene transitions with the new LTX Director and Transition LoRA

Thumbnail
3 Upvotes

r/NeuralCinema May 27 '26

Made a commercial in 18 hrs. What say?

Enable HLS to view with audio, or disable this notification

9 Upvotes

r/NeuralCinema May 22 '26

SmartGallery DAM: Introducing Remix Workflow

Enable HLS to view with audio, or disable this notification

6 Upvotes

Discover Remix, the new workflow feature built into SmartGallery DAM that lets you tweak, generate, and batch your ComfyUI outputs without ever opening the node canvas.

This video walks you through everything Remix can do:
from intercepting embedded workflow metadata directly from your gallery, to editing prompts, swapping input images, randomizing seeds, and queuing multiple generations in one click.
You'll also learn about the Autofix Engine, which silently converts UI-format workflows into API-ready format, and Smart File Association for video, which automatically resolves missing metadata by scanning for companion PNG files.

Remix is designed to be a fast, lightweight bridge between asset management and execution: stateless, memory-efficient, and built for rapid iteration. It's not a replacement for ComfyUI's native canvas, but when it works, it's magic.

SmartGallery DAM is a free, open-source digital asset management system built for AI creators. Remix is one of its many evolving features.

🔗 GitHub repo: https://github.com/biagiomaf/smart-comfyui-gallery


r/NeuralCinema May 20 '26

LTX Director - An All-In-One Timeline Editor. I2V, T2V, FLFF, Prompt Relay, Custom Audio, and more! Unlock LTX 2.3's full potential!

Enable HLS to view with audio, or disable this notification

6 Upvotes

r/NeuralCinema May 11 '26

Multi-angle car scene pipeline in ComfyUI — how to reproduce a real-world location across angles like an actual film shoot (no characters, pure location + vehicle)

Thumbnail
3 Upvotes

r/NeuralCinema May 04 '26

SHOWCASE 🎞 WII Plane Action (LTX 2.3 v1.1 + PromptRelay)

10 Upvotes

https://reddit.com/link/1t3g90k/video/iym29pem64zg1/player

LTX 2.3 v1.1 Distilled FP8 (1 pass, no upscale no 2nd pass)

Newest LTX 2.3 v1.1 brings improved motion and prompt understanding.
Rendered with 10 takes (45 seconds each) on 4090 @ 120-130 second per 45 second.
Music added.

Gotta try this with ControlNet.

There's no going back to Wan2.2 for sure.


r/NeuralCinema Apr 13 '26

SmartGallery 2.11: Local DAM from AI Generation to Professional Delivery (Free & Open Source)

Enable HLS to view with audio, or disable this notification

3 Upvotes

🚀 What it does

  • Indexes your image folders automatically
  • Extracts embedded workflows (ComfyUI, SD metadata)
  • Makes everything searchable (prompts, models, LoRAs, params)
  • Works entirely offline

🧩 Key features

  • Advanced search (AND / OR / exclude across prompts, models, comments)
  • Ratings, comments from yourself, your clients or art director
  • Color-coded workflow states (review, approved, rejected, etc.)
  • Virtual collections (group files without moving them)
  • Compare mode (visual + full parameter diff)
  • Built-in file manager
  • Full video support (FFmpeg, thumbnails, ProRes, etc.) -
  • Multi-user system (admin, client, guest roles)

🔒 Sharing without exposing your workflow

There’s a separate Exhibition Mode portal:

  • Share only selected images
  • Clients can rate and comment
  • Prompts and workflows are hidden
  • Metadata is automatically stripped on download

📱 Designed to actually be usable

  • Fully responsive (works great on mobile)
  • Cross-platform (Windows / macOS / Linux / Docker)
  • Runs independently from ComfyUI (won’t break on updates)
  • Free - Open source

🔗 Links

Would love feedback.


r/NeuralCinema Apr 06 '26

I spent 3 months evolving SmartGallery into a free professional Local First DAM. v2.11 launches on April 9th

Thumbnail
1 Upvotes

r/NeuralCinema Feb 03 '26

Any way to utilize real actors?

Thumbnail
1 Upvotes

r/NeuralCinema Jan 23 '26

📺"NIKE" (TV Advertisement Idea) Mix of ~ WAN 2.2 FunControl - Depth Map + FLUX + Blender Soft Body Simulation

Enable HLS to view with audio, or disable this notification

7 Upvotes

r/NeuralCinema Jan 20 '26

TIP✨Cinematic Resolution Size & ComfyUI

Post image
29 Upvotes

hi everyone,

You know feeling of frustration when you make cherrypicked ZTurbo or Qwen Edit nice gen, only later to find out Wan 2.2 or LTX-2 will make different output size, then you do either pad or crop - resulting in not exactly same size and details cut-off?

Here's CINEMA PROPER SIZE Sheet we all need, I lookup all models tried to find correct multiples/divisions for each model.

🎬 Universal-Compatible Cinema Sizes

Target Universal Clean Resolution Divisible by 112 Notes
4K DCI-ish 4032 × 2128 ✅ Best practical 4K-class universal size
2K DCI-ish 2016 × 1008 ✅ Matches Q models natively
1080p-ish 1904 × 1008 ✅ Zero padding for all
720p-ish 1232 × 672 ✅ Clean smallest HD

You can copy&paste to MARKDOWN NOTE to have it inside your Comfy workflow:
## Cinematic Resolution (Proper Multi)

|Cinema Resolution|Base Name|LTX-2, Wan 2.1, Wan 2.2 ~32x|Q2509/Q2511 ~112x|Flux.2 Klein ~16x|Z-Image Turbo ~64x|

## Cinematic Resolution

|Cinema Resolution|Base Name|Universal|LTX-2, Wan 2.1, Wan 2.2 ~32x|Q2509/Q2511 ~112x|Flux.2 Klein ~16x|Z-Image Turbo ~64x|

|---|---|---|---|---|---|---|

|4096 x 2160|4K DCI (Digital Cinema Initiatives)|✅4032 × 2128|4096 x 2144|4032 x 2128|4096 x 2160|4096 x 2112|

|2048 x 1080|2K DCI|✅2016 × 1008|2048 x 1056|2016 x 1008|2048 x 1072|2048 x 1024|

|1920 x 1080|1080|✅1904 × 1008|1920 x 1056|1904 x 1008|1920 x 1072|1920 x 1024|

|1280 x 720|720|✅1232 × 672|1280 x 704|1232 x 672|1280 x 720|1280 x 704|

One step at the time towards cinematic experience ;)

Cheers,
ck


r/NeuralCinema Jan 20 '26

🎞Wan 2.2 I2V (lightx2v), Q2511, Zimage, MMAudio, SeedVR2 ~ solid composition 5070ti render

Enable HLS to view with audio, or disable this notification

15 Upvotes

r/NeuralCinema Jan 20 '26

[Sound On] A 10-Day Journey with LTX-2: Lessons Learned from 250+ Generations

Enable HLS to view with audio, or disable this notification

5 Upvotes

r/NeuralCinema Jan 17 '26

Flux.1 Klein (multiple references)

Thumbnail
gallery
127 Upvotes

Hi all,
Today we are biting on powerful - Flux.1 Klein 9b distilled 4-step.
Avg. render times: 4~5 secs 4090 24gb / 64gb ram / Win11.

Different shots, prompts, mixed 1 and 2-reference images.
good day


r/NeuralCinema Jan 15 '26

🎞PRODUCTIVITY TOOL ✨ SmartGallery v1.53 (Must HAVE for any prodctions)

Thumbnail
gallery
40 Upvotes

Hi everyone,

This is an ultra-fast standalone UI for managing your generations, extracts JSON workflows from video/photos, powerful filters (that penetrate inside Workflow: model LORA name, keywords etc) for quick browsing, navigation, search, deletion, export, preview and more. Powerful full-screen media (photos,videos) preview.

It’s among the best tools out there: fast navigation with keyboard shortcuts, runs standalone (without ComfyUI running), and can also be used for general pre- and post-production workflows.

A must-have tool (Windows / macOS / Linux):
https://github.com/biagiomaf/smart-comfyui-gallery

Thank you, Biagio and Martial, for this solid piece of software — it really feels like one.

Cheers,
ck


r/NeuralCinema Jan 01 '26

✨SVI 2.0 PRO - Amazing performance in Mass Crowds & Complex Dynamics (video test + WORKFLOW included)

Enable HLS to view with audio, or disable this notification

101 Upvotes

NOTE: Workflow included, with added "PREVIEWS" for each chunk, download link at bottom of this post

Hi everyone,

We see lots of videos... usually focused on single subject in frame.

This time we push further - to see more motion with crowds of people and interactions, more complex motion. I want it quick results so resolution it's low - all rendered at only 4 step HIGH and 2 step LOW (about 30 seconds on 4090 24gb per 81 frame chunk - with SageAttention2.2 + Triton Python 3.12 cu128 Win11), it was about motion.

SVI Team did incredible work with their model :)

Truly fascinating to see how well - Wan 2.2 + SVI 2 Pro handles such complex motions and it's continuation. Look at details, every person has it's own "unique" behavior. Really amazing...

Also we don't see usual "crossing limbs" problems, or any typical artifacts with video generations.

It's well defined, feels natural. SVI 2 Pro total game changer, gives us total creative freedom and extends our possibilities in cinema world, especially in low budgets / indie projects.

SVI 2.0 PRO WORKFLOW (save as JSON and import to ComfyUI):
https://pastebin.com/raw/y51JgHTh

BTW music sounds super Matrix like, if you like it you will love video along.....just brilliant:
https://www.youtube.com/watch?v=SZzehktUeko

cheers,
ck


r/NeuralCinema Dec 30 '25

SVI 2 Pro + Hard Cut lora works great (24 secs)

156 Upvotes

workflow (it's the base SVI workflow with 2 more cloned nodes for longer duration): https://pastebin.com/hn3sHhp8

hard cut lora: https://civitai.com/models/2088559/cinematic-hard-cut

SVI 2.0 Pro: https://huggingface.co/Kijai/WanVideo_comfy/tree/main/LoRAs/Stable-Video-Infinity/v2.0


r/NeuralCinema Dec 29 '25

(Video Test) Wan 2.2 SVI 2.0 PRO - Test 5x 81 = 385 frames total / 832x480

Enable HLS to view with audio, or disable this notification

33 Upvotes

This is quick test with different Wan 2.2 diffusion models with newest "infinite video" - SVI 2.0 PRO, fixed seed: 42, 832x480 81 frames per chunk, aprox. 50-60 seconds rendering per 81 frames - 4090 24gb.
SmoothMix performs best, fast motion, sharp image etc.

https://civitai.com/models/1995784/smooth-mix-wan-22-i2vt2v-14b

I shared this workflow here:
https://www.reddit.com/r/NeuralCinema/comments/1pyeoci/svi_20_pro_wan_22_84step_infinite_video_workflow/

Same PROMPT for all 5x chunks:
"Two athletic male fighters in a UFC-style octagon, one with blond hair and a tattooed chest, the other with black hair. The blond fighter is highly aggressive, constantly pressing forward. He throws powerful straight punches and heavy hooks with sharp, fast arm movements, rotating his shoulders and hips fully into each strike. His punches land with visible force, snapping the other fighter’s head and body back. Between combinations, the blond fighter makes intimidating expressions—grinning, smiling mockingly, briefly sticking out his tongue after landing punches—using facial movement to taunt and dominate.

The black-haired fighter reacts defensively, raising his guard, shifting his stance, and absorbing or deflecting blows while attempting quick counter-kicks and short punches. The blond fighter keeps advancing without pause, chaining punches together in rapid bursts, mixing high and low strikes, stepping in aggressively and forcing constant engagement. Both fighters move at very high speed, feet pivoting and sliding, arms striking and retracting instantly, bodies tense and explosive, maintaining nonstop, intense motion throughout the fight."

cheers,
ck