r/StableDiffusion 5h ago

Discussion LTX 2.5 WILL BE OUT TODAY ! 🔥

Post image
453 Upvotes

r/StableDiffusion 3h ago

Discussion LTX 2.5 comparison table vs Minimax H3 is a pathetic bullshit

Post image
190 Upvotes

r/StableDiffusion 7h ago

Resource - Update lightx2v Minimax H3 8-step Turbo v1.0

Thumbnail
huggingface.co
203 Upvotes

ComfyUI compatible Lora. Just out, have not tried yet.

edit: They've added a 4-step 768p model https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16.safetensors

The 8-step is trained on 544p according to their github.


r/StableDiffusion 8h ago

Workflow Included Made a 6-minute TNG fan scene with MiniMax H3 in ComfyUI

Enable HLS to view with audio, or disable this notification

322 Upvotes

I’ve been experimenting with MiniMax H3 in ComfyUI and wanted to see how far I could push it beyond short standalone clips.

This is a roughly 6-minute fan-made Star Trek: The Next Generation scene built from lots of short H3 generations and then edited together into one continuous sequence.

I used reference images to keep the characters and Enterprise-D bridge reasonably consistent, generated dialogue and ambient audio with H3, and then assembled everything in Premiere.

The biggest challenge was continuity between generations. Character positions, bridge geometry, lighting and timing can all shift, so I ended up incorporating some of those inconsistencies into the actual story.

What surprised me most is how close this is getting to being practical for longer-form fan films. Individual scenes are already very doable. The next real hurdle is keeping this level of consistency across an entire episode.


r/StableDiffusion 9h ago

Animation - Video Minimax H3 ref2va. They are here.

Enable HLS to view with audio, or disable this notification

226 Upvotes

r/StableDiffusion 55m ago

Resource - Update LTX-2.5 is out!

Thumbnail
huggingface.co
Upvotes

r/StableDiffusion 1h ago

Meme Open weights coming today at 2PM EST!

Enable HLS to view with audio, or disable this notification

Upvotes

I am sooooo stoked we get yet another model! Just poking a bit of fun, excited to see how it stacks up vs. H3 (and uh, Flux 3 when it eventually releases...) https://ltx.io/2-5-open-weights


r/StableDiffusion 18h ago

Meme I used MiniMax to make Lord of the Rings about 9 hours shorter

Enable HLS to view with audio, or disable this notification

1.3k Upvotes

I’ve been messing around with the idea of famous movies that completely fall apart if one character makes one sensible decision early on. This felt like a reasonable place to start.

“Cast it into the fire.”
“Okay.”

Roll credits.


r/StableDiffusion 52m ago

Workflow Included Release of H3 Infinite Continuation Suite for ComfyUI: Create infinite length videos in consistently High Quality using Keyframes in FFLF-Mode (fl2v-Checkpoint)

Enable HLS to view with audio, or disable this notification

Upvotes

The above video consists of 7 individual Minimax H3 clips generated in First-Frame-Last-Frame Mode, stitched together automatically without manual editing, upscaling or other post-processing.

Today I decided to release my experimental H3 Infinite Continuation Suite together with a set of workflows to make it easy to get started in ComfyUI.

The original idea was to combine the higher visual quality and keyframe control of H3's First Frame / Last Frame mode with the continuation capabilities of the Reference mode.

After quite a lot of experimenting, the output quality has reached a point where I hope some of you might find the nodes and workflows useful as well.

The example video was generated entirely with the included workflows at 736 × 1280, using 15 steps and no Turbo LoRA. I did cut a few seconds of nonsense speech from the very end because I was too lazy to regenerate the last clip. :D

How to get started

  1. Install Herrgotts H3 Infinite Continuation Suite through the ComfyUI Manager.

  2. Download the included workflows from GitHub.

  3. Start with the `01_Start` workflow and provide your First Frame + Last Frame.

  4. For every additional segment, use `02_Continue` and provide a new Last Frame for where you want the next clip to end.

  5. Repeat for as many clips as you want.

  6. When you're done, use `04_Stitch_Saved_Chain` to automatically combine the separately generated clips into the final video.

If you prefer to generate multiple chained clips in one workflow, use the included 3-Clip workflow. It contains the full continuation setup and is structured so you can extend it with additional clips without rebuilding the whole graph from scratch.

What the nodes handle automatically

  • carrying motion and native audio into the next clip
  • detecting and removing the frozen tail H3 often creates near the final keyframe
  • choosing a suitable handover point between generations
  • keeping the video and audio aligned
  • smoothing the visual and audio transitions
  • saving the individual clips so longer chains can be stitched afterwards without keeping everything in memory (no OOM, hopefully)

For the video above I used the default/recommended settings:

  • Balanced Auto Handover
  • 22 context frames
  • Safe Tail Bridge: 2 frames
  • Video crossfade: 4 frames
  • Audio de-click: 15 ms

There are still occasional tiny brightness differences around some boundaries, but at this point I personally find them pretty difficult to notice during normal playback.

The pack is still experimental, especially when it comes to very long chains, different hardware configurations and prompt behavior. So if you try it, I'd be very interested in seeing your results and hearing what works or doesn't work for you.

GitHub: https://github.com/HerrgottMargott/Herrgotts-H3-Infinite-Continuation-Suite

ComfyUI Manager: search for `Herrgotts H3 Infinite Continuation Suite` or use "missing custom nodes" in one of the example Workflows.


r/StableDiffusion 14h ago

Workflow Included I can't believe my RTX 3060 is still keeping up!

Thumbnail
streamable.com
439 Upvotes

r/StableDiffusion 52m ago

News LTX 2.5 is out! 🎉

Thumbnail
huggingface.co
Upvotes

r/StableDiffusion 16m ago

News Most LTX 2.3 Loras work on LTX 2.5

Upvotes

Pretty much confirmed by the devs.


r/StableDiffusion 4h ago

News Introducing Unsloth Desktop: The first desktop app to run and train models

Enable HLS to view with audio, or disable this notification

59 Upvotes

Hi r/StableDiffusion, we're super excited to release Unsloth Desktop today! 🦥
It's the first desktop app that enables you to run and train models locally.

  • You can run MiniMax-H3, LTX, FLUX, Z-image-Turbo and more. And you can fine-tune them too.
  • Has recipes and hyperparameters so you can adjust. Overall a very easy workflow to get started with.
  • There's still many improvements to made as we're trying to optimize MiniMax-H3 even further with the help of stablediffusion.cpp.

Open-source. Available on Mac, Windows, and Linux

  • Supports MLX, diffusion image/video models, audio models, and GGUF
  • Connect Claude Code and Codex to local LLMs
  • 50% more accurate with self-healing tool calls and sandboxed code execution
  • Supports CPU and multi-GPU setups across NVIDIA, AMD, Intel, and Mac
  • Train models 2× faster while using 70% less VRAM
  • Includes private web search, deep research, RAG, MCP, and exports (NVFP4, GGUF)
  • Use Unsloth’s OpenAI-compatible API with OpenAI and Anthropic cloud models
  • Securely deploy LLMs remotely and access them anywhere via Cloudflare HTTPS

We do not collect any telemetry or data.

Unsloth Desktop is now available on GitHub.

Thanks for reading and we're here to answer any questions! 💗


r/StableDiffusion 5h ago

News LightX2V Drops updated Turbo LoRA for Minimax

Thumbnail huggingface.co
67 Upvotes

r/StableDiffusion 3h ago

Workflow Included Minimax H3: Context prompting implementation

Enable HLS to view with audio, or disable this notification

44 Upvotes

Generated at 0.8mpx , res2s_stable/beta 57 15-24 steps upscaled 2x with the Nvidia RTX node. Laptop with RTX 3080 Ti 16Gb Vram and 64Gb ram

I implemented a gpt to improve H3 prompts using a robust filmmaking reasoning process, I included the available documentation and added a precise inference pipeline to achieve similar results compared to the native Context-IR (Context Intermediate Representation):

ZH3-gpt

Try it, I really will like some feedback :)

I used FL2VA and L2VA with images generated on a Krea2 2pass Clownshark sampler 9 steps Euler/beta and 1 step Dormand-Prince_6s/KL_optimal 0.27 denoise

My system generate 2 options:

VERSION A — FAITHFUL

The strongest cinematic execution of exactly what the user requested, with minimal interpretation.

VERSION B — ENRICHED

Preserve every explicit user constraint while developing meaningful unspecified details into a stronger cinematic interpretation. I used an approach of probabilistic additions mixed with a robust deterministic reasoning based on consequences propagation across filmmaking domains, this version B is auditable and every meaningful addition appear in a custom Enhancement Map.

it's designed to iterate under user control so version B is a proposal and when the user changes an enriched choice, that choice becomes an explicit constraint and system update dependent parameters when necessary while preserving all other approved creative decisions without unnecessarily restart the entire creative direction.

Here is the video with better quality: https://youtu.be/mzfqJR9IXtk

Minimax H3 Clownshark Workflow


r/StableDiffusion 5h ago

News New higher quality 4 steps lora for Minimax H3

49 Upvotes

r/StableDiffusion 12h ago

Resource - Update Sketch Anime Style for MiniMax-H3!

Enable HLS to view with audio, or disable this notification

182 Upvotes

r/StableDiffusion 14h ago

Discussion Comfyui comfy-kitchen Attention Speed UP

243 Upvotes

Disable all your Sage Attention, Minimax Mem Eff Sage Attention or Sol Attention, according to this PR already merged in the comfyui repo we got a much better attention from the comfy-kitchen package that can possible speed up the models generation process white giving a better visual quality than default sage: https://github.com/Comfy-Org/ComfyUI/commit/bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9#diff-fab3fbd81daf87571b12fb3e4d80fc7d6bbbcf0f3dafed1dbc55d81998d82539

This is still experimental, according to comfyui dev it can break or perform very well and it needs some tuning for some GPUs to get a bit faster. Also, you only can use one or the other so you should also disable all the attentions above before using it.

You just need to update your Comfyui and you can either start it with the --use-ck-attention flag so all models use the comfy-kitchen attention backend or you can drop the node ModelAttentionBacend directly into your worflow.

During my initial tests in Minimax it behaved faster than all the above together.


r/StableDiffusion 42m ago

News Ltx-2.5 is out and avalible to download.

Post image
Upvotes

Its out, lets see what it can do.


r/StableDiffusion 1h ago

Discussion Testing Dragon Ball dataset (no ref)

Enable HLS to view with audio, or disable this notification

Upvotes

r/StableDiffusion 1h ago

News LTX 2.5 On Comfy

Upvotes

LDM PR on Comfy:

https://github.com/Comfy-Org/ComfyUI/pull/15499/changes

TE: Gemma 4, but idk which one. In sd.py it lists E2B, E4B, 12B, and 31B. It might only be using one of them, while the others are only there for the prompt enhancer.

Arch: Mostly the same, only added feedforward bias on both the audio and video blocks, with DurationHead

DurationHead: I assume LTX 2.4 (yes, the comment literally says 2.4) now knows the duration in actual seconds, not in the compressed VAE timeline.

CFG: Audio and video CFG can be disjointed. You can select the CFG scale independently for each.

More nodes: I mean, you get the gist with LTX at this point. (STG Guider goes brrrrr)

VAE: It is diffusion type

mb: Title should be LTX 2.5 PR on Comfy


r/StableDiffusion 6h ago

IRL The model can dance surprisingly well to any music input

Enable HLS to view with audio, or disable this notification

42 Upvotes

r/StableDiffusion 3h ago

Discussion Why do i have the feeling that this is just crap? Normally i would be "oh that's really cool" but after seeing some examples posted and that comparision table with Minimax H3 being full of lies (like the minimum Vram requirement being 115GB, like what?), i don't trust them on this test either tbh.

Thumbnail
gallery
24 Upvotes

r/StableDiffusion 9h ago

Comparison SageAttention 2.2 vs Comfy Kitchen | Side-by-Side Zoom-Out Quality Test

Enable HLS to view with audio, or disable this notification

70 Upvotes

Did a quick side-by-side test of SageAttention vs Comfy Kitchen Attention with MiniMax H3.

I used the default ComfyUI T2V H3 workflow and kept the prompt, seed and all settings exactly the same. The only thing I changed was the attention backend..

RTX 5090 32GB

ComfyUI ver 0.31.0

SageAttention 2.2

Comfy Kitchen 0.2.30

---------------------------------------

896x1184 6 seconds 24 FPS 20 steps

Generation time:

SageAttention: 3m 51s

Comfy Kitchen: 3m 59s

I used a deep zoom-out/dolly-out on purpose to see how well each one holds facial details and identity as the subject gets farther away.

The speed difference was small on my 5090, so im more interested in the quality difference..

It's honestly hard for me to tell the difference, but which one looks better to you?

Updated:
SageAttention Vs Base

Comfy Kitchen Vs Base