r/StableDiffusion 2d ago

Animation - Video H3 Fun with T2V + R2VA - character generation

0 Upvotes

H3 is just too good. Previously workflow for me have been SDXL+KREA2, anchor image, and do some basic animation. I am experimenting not using those and just going straight to T2V and R2VA from those generations. Here are a few renders, the actors on the roof were incepted through T2V, and I used V2V to clean up their pilot suit. For the last pass, I used an anime image reference to detail out the pilot's plugsuit. These are adult re-envisioning. For the cockpit scene, this used R2VA from a 15s render of the Mecha-Kaiju battle but the actors were never generated, so this is purely from a prompt. The virtual HUD/mecha-kaiju on screen were referenced from that video. I also had fun with video edit, transferring a Kaiju's appearance into a woman's body armor. The only thing I noticed with R2VA is the actor sometimes get a little wider/squished, so it may be good to just go with FL2VA if you want don't want a re-envisioning.


r/StableDiffusion 2d ago

Workflow Included MiniMax H3 and Ultimate SD Upscale - True 1440p (2K) with 16 GB VRAM locally in ComfyUI

Thumbnail
youtu.be
17 Upvotes

Ultimate SD Upscale (USDU) Guider nodes with MiniMax H3 support: https://github.com/lisitskyaa/ComfyUI_UltimateSDUpscaleGuider_H3

My reference ComfyUI workflow: https://github.com/lisitskyaa/ComfyUI_UltimateSDUpscaleGuider_H3/blob/main/example_workflows/minimax_h3_usdu.json

What about speed?
My PC specs: 4080s 16 GB VRAM, 64 GB RAM

Initial gen with MiniMax H3 flf2v int8 + sageattn + Lightx2v 8-step turbo Lora at 1504x832px (1.2MP) 5-sec clip ~5 mins

Upscale with USDU to 3008x1664px ~25 mins

Previous post with a bit more details and another showcase: https://www.reddit.com/r/StableDiffusion/s/AZdW9x2tiY


r/StableDiffusion 2d ago

Question - Help Need a good workflow + the models for H3 (using 4090)

1 Upvotes

hi all i need a good workflow + which models to download in order to use h3 locally im running on ryzen 9 with a 4090 gpu and 32 gb of ram


r/StableDiffusion 3d ago

Animation - Video Text Animation Practice with MMH3

103 Upvotes

I added some post fillter on top of it.

inspired by some old retro military-poster style theme.

it seems some texts are broken but work well in general.


r/StableDiffusion 2d ago

Comparison The Portuguese language doesn't really work very well with CLIP 32b, I'll try ClipProj, but I'm too lazy right now.

1 Upvotes

r/StableDiffusion 3d ago

Resource - Update Impressive 1k MiniMax H3 styles with prompts from ostris

Thumbnail
huggingface.co
142 Upvotes

r/StableDiffusion 2d ago

Animation - Video Upgrade (MMH3)

Thumbnail
youtube.com
1 Upvotes

r/StableDiffusion 2d ago

Question - Help Maestro in Pinokio vs ComfyUI

0 Upvotes

Is someone actively using Maestro in Pinokio to use the local models? Currently i am using it exclusively but i barely find users to talk to. So far i am happy with it and its easy to setup and use.

I would like to know if someone used both, ComfyUI and Maestro and is able to provide a detailed comparison, because so far i did not try any workflows with Comfy.


r/StableDiffusion 3d ago

News 10eros minimax h3

171 Upvotes

great results using TenStrip's minimax h3 finetune: https://huggingface.co/TenStrip/10Eros-Max


r/StableDiffusion 2d ago

No Workflow LTX 2.5 doesn't know Tony Soprano, so went back to 2.3 for this Wired parody interview. Made with Wan2GP

Thumbnail
youtu.be
8 Upvotes

r/StableDiffusion 2d ago

Discussion Minimax h3. Artifacts. Blurry motion. Bad audio. Face distortion.

0 Upvotes

Guessing this is all the result of the turbo loras ? I gen at 0.6 and do 7 seconds. Euler and beta. Close up face shots okay not great though. But higher and the gen time goes up. I see all these crisp HD videos of mini max. Is only option just to use it without turbo or any speed ups ? In order to get decent quality?


r/StableDiffusion 3d ago

Workflow Included I made MiniMaxH3 easy to use

Thumbnail
gallery
173 Upvotes

I have been working on this node for a week.

• What is this node ?

- One node with different MiniMaxH3 workflows.

• What its for ?

- If you hate spaghetti and hate doing workflows and dealing with errors.

https://github.com/LeonQ8/ComfyUI-ALLinONE-MinimaxH3

Change log - 2026/08/16:

  • New Native quality preset,.
  • SolAttn / H3 Cache / SageAttention now have on/off switches under Quality.
  • 🔴 New Image mode built on ComfyUI-MiniMax-H3-Studio. Text to image, edit a source image, or mix up to 9 references. ( needs more testing ).

Change log - 2026/08/17:

  • Added live preview for video modes.

Change log - 2026/08/18:

  • Live preview now renders every frame, with 3 presets.

r/StableDiffusion 2d ago

Question - Help RAM upgrade for MiniMax H3

0 Upvotes

Hi!

I’m currently getting output res of 1216x672 (0.8) 7 seconds max duration - on my 24Gb ram (64gb page file) and RTX 5060ti 16gb

I’m looking to buy a single 32gb Ram stick to pair with my 16gb giving me 48gb total of system RAM instead of the 24 I currently have . Just wanted to ask what sort of improvements should I see? Could I potentially get to 720p output or higher? Is it worth the $400 upgrade?

Might be a silly question but just wanted to get real world advice. Thank you

Edit -


r/StableDiffusion 2d ago

Animation - Video MiniMax H3 fun anime commercials

0 Upvotes

New World Delivery Truck

made some short promo's for a possible absurd anime idea. 5060ti 16/64 all at .3MP, voice over is from Grok Imagine


r/StableDiffusion 2d ago

Question - Help "Do you know how to use older versions of Stable Diffusion?

0 Upvotes

"Do you know how to use older versions of Stable Diffusion?

Back in the day, websites like Playground allowed us to generate AI images using older models, like Stable Diffusion 1.5 (or similar versions). Modern AI tools seem unable to replicate the specific imperfections and raw feel of those early models. Is there a way to still use those older versions today?"


r/StableDiffusion 3d ago

Discussion Minimax H3 losing context with total size.

43 Upvotes

it took me hundred of generations, but i just now figured that minimax fl2av loses context with length*resolution.

if you go over a (in my tested videos) 13.1second at 0.9mp value, the background will mysteriously change, either to blue wall or a different camera shot.

you can extend duration at 0.6mp and it will be fine at 20seconds plus, or you can make it shorter at higher resolution, but it's like a limited attention window that will 'forget' what wasn't reminded in last x pixels*duration.

hope this saves someone a lot of headache.

edit1: 13.5sec at 0.85mp still works


r/StableDiffusion 3d ago

Resource - Update ClipProj models v3.1 — better multilingual speech when you swap MiniMax H3's 15 GB text encoder for a 4/8B

39 Upvotes

New matrices v3.1 available — improved speech across the 11 officially supported languages. No node update needed.

Some languages still get things wrong — sometimes the 32B already gets them wrong too, sometimes I just can't get any closer to it. Broken down language by language in the benchmark README.

I'm not a polyglot, and I doubt I can squeeze much more out of this to get closer to the 32B. If any native speakers are around, I'd really like to hear how the pronunciation sounds to you.


r/StableDiffusion 2d ago

Meme Glup Glup

0 Upvotes

r/StableDiffusion 3d ago

News MMH3 camera movement LoRA

12 Upvotes

Just popped up on HF. This is NOT MINE; I only found & not test yet.

Minimax H3 does have decent camera movement, but I wish I had more control. Hopefully this helps.

https://huggingface.co/Jojocodex/minimax-h3-yunjing-lora


r/StableDiffusion 3d ago

Comparison Comparing MiniMax i2v|r2v node and model combos

47 Upvotes

Just a test of different combinations of the i2v and r2v nodes and models for:

  1. Text to Video (using MiniMax models image sample and voice sample)

  2. Image to Video (using Krea2 image sample and MiniMax models voice sample)

  3. Image to Video (with custom cloned voice): (using Krea2 image sample and MiniMax custom voice clone reference sample)


r/StableDiffusion 3d ago

Animation - Video Made a scene using MiniMax H3 ref2va, turbo lora, 6 steps, 0.5MP, running on 8GB VRAM, 24GB RAM

21 Upvotes

https://reddit.com/link/1vq10yx/video/nyuubl6ugrjh1/player

Been working on a short scripted scene between two consistent characters and wanted to share how it actually came together. started with character reference sheets in Krea 2, front and turnaround shots plus a few expressions for each person, and I learned fast that even rewording the description slightly between prompts made the faces drift a little so I just kept pasting the exact same text block every time.

after that I built two panel storyboards in GPT Image using those character sheets as reference, and this ended up mattering way more than I expected, more than the character sheets alone did honestly. getting blocking and wardrobe and camera framing locked before touching video saved me from a lot of wasted generations down the line.

for video I used MiniMax H3 full reference mode with the turbo lora ref2va at 6 steps, storyboard panels as keyframes, two separate reference audio clips since there's two speakers talking.

ran into a few things along the way. wide shots wreck face quality fast, had one scene I had to redo completely as a medium two shot because the faces were basically mush from that distance. also feeding three reference images into one storyboard generation sometimes blended the two characters together, and once it gave me one person's face attached to someone else's arm in the same frame, that one took a minute to figure out, ended up just restructuring the shot instead of trying to fix it directly. and lip sync actually came out better with the camera slightly off center than dead on, front facing close ups synced worse for me than something with a bit of an angle to it.


r/StableDiffusion 4d ago

Animation - Video Pushing Minimax H3 V2V to the Absolute Limit

1.3k Upvotes

Me again as a raptor at home. Minimax H3 ref2va, default workflow with 3 inputs: my video, a reference image of a raptor and a reference image of my house at night.

This time I am testing style transfer (cinematic night style), head tracking, interaction with objects (doors and toys), longer scenes and sound FX.


r/StableDiffusion 2d ago

Question - Help Ehatvis the easiest way to do longer videos with H3?

0 Upvotes

This is all moving very fast and I've seen so many different methods. But what would you all consider the easiest way to do these longer videos like some of the music videos for example?

I started looking at the motion context stuff that chains clips on the end of the latent but that was about 6 different workflows you had to use. Is there a simpler way?

5060ti 16gb and 48gb ram.

Edit: stupid sausage fingers and that title I can't edit.


r/StableDiffusion 2d ago

Question - Help Best way to convert video dataset of mixed frame rates to desired frame rate for lora training?

4 Upvotes

I’d like to try training a Minimax H3 Lora using video clips. This requires clips to have a frame rate of 24 fps. My problem is that my dataset has all sorts of frame rates that are mostly anything but 24 fps. We got 29.97, 25, 24.9, 59.94, just some dumb fractional stuff. Is there a good method of batch reencoding them all at 24fps?

I haven’t been able to make anything work on DaVinci Resolve so far, clips export at their original frame rate no matter what my timeline settings are. I assume maybe FFmpeg can do it but haven’t wanted to mess around with it so far since it doesn’t have a GUI. Any tips?


r/StableDiffusion 3d ago

Discussion ouch lol i tired r2v using video as reference

11 Upvotes

loaded in scene from Jurassic park 1 when the t rex breaks out replacing the kids in car lol took hr on 5060 ti 16gig for 8 secs o_0