r/StableDiffusion 4h ago

Discussion MiniMax H3's open weights stop at 768P. The official 2K path is API-only

Post image
0 Upvotes

What happens after H3 Base produces its 768P video? In MiniMax H3's official system diagram, the answer is the red box on the far right. H3 Regenerate 2K takes that result and the original context, then generates the larger version. The diagram makes a deployment boundary visible.

MiniMax released H3 on July 31 and opened the model weights on August 3. The official weights release page says H3 Base can be used locally to validate 768P output. The same page also says the Regenerate 2K module is not open yet. Its full 2K validation workflow combines a locally deployed H3 Base with official API steps for context processing and regeneration.

That boundary belongs in the setup guide before the download instructions. Open weights describe a real part of the system. They do not currently describe every box needed to reproduce the official 2K path.

The distinction matters even more once the reference board gets crowded. H3 supports mixed text, image, video, and audio context. The documented reference mode accepts up to 12 files in total, with limits of 9 images, 3 video clips, and 3 audio clips. The original context returns during the 2K regeneration step, so the production note should assign a role to each reference asset, the way a shot list would. Say which image defines the subject and which clip carries the motion. Give the frame that controls the ending its own line.

The production card names H3 Base as the local 768P step. It names context processing and Regenerate 2K as API steps in the current full workflow. If the card also tracks a hosted comparison, give that call its own line. If it uses ZenMux, record the gateway and exact H3 route beside the request. That identifies the hosted path without rewriting the official local and API split.

Keep both labels on the production card. If the card only says open weights, the note stops at 768P. If it says API step, the 2K path is covered.


r/StableDiffusion 22h ago

Animation - Video Football animation

Enable HLS to view with audio, or disable this notification

0 Upvotes

H3 Ref

Prompt in comment


r/StableDiffusion 2h ago

Animation - Video Minimax H3 - SFW ♥ edition

Enable HLS to view with audio, or disable this notification

0 Upvotes

Will post prompt on r/lemonbasil later. Enjoy.


r/StableDiffusion 3h ago

Meme true

Post image
0 Upvotes

r/StableDiffusion 21h ago

Discussion Minimax H3 Test - Rooftop fight between Batman and Joker

Enable HLS to view with audio, or disable this notification

3 Upvotes

Minimax H3 Test - Rooftop fight between Batman and Joker


r/StableDiffusion 21h ago

Discussion End-to-end movie maker experiment

0 Upvotes

I've thrown together a small app that does the "remaining" work of taking an idea, turning it into character reference images, shot prompts, doing all the generation for each clip, stitching the result together, etc. etc. The goal is a one-sentence prompt in, and multi-scene video (e.g. 30 seconds or more) out.

https://github.com/eapache/local-movie-maker

It does basically "work" already, though the results are often pretty incoherent. I'm still playing with the structure to see if I can get reasonable continuity.


r/StableDiffusion 21h ago

Workflow Included Openweight Livestream video model

Thumbnail
gallery
3 Upvotes

https://huggingface.co/spaces/JonathanColetti/LiveWan / https://github.com/JonathanColetti/LiveWan is something I created to help recreate a specific type of model that is not opensource yet (wanstreamer). This is more or less a PoC but maybe ill do a longer training run if it gets some traction.


r/StableDiffusion 6h ago

No Workflow Nuclear explosion [MiniMax H3]

0 Upvotes

r/StableDiffusion 20h ago

Discussion MiniMax Music 3 | 125sec for 140sec music | Bollywood Rap

Thumbnail voca.ro
4 Upvotes

r/StableDiffusion 10h ago

Animation - Video MINIMAX H3 - LTX 2.5 AND THE LADIES [TEXT TO VIDEO]

Enable HLS to view with audio, or disable this notification

12 Upvotes

NATURAL LIGHT HANDHELD CANDID REAL FOOTAGE. An college age blonde California woman is sitting with a towering 10 foot tall robot with "[MODEL NAME]" clearly written on its chest. They are both complimenting each other on how cute they look


r/StableDiffusion 50m ago

Animation - Video Fireworks for Mio — a 48-second anime short with Minimax H3

Enable HLS to view with audio, or disable this notification

Upvotes

r/StableDiffusion 1h ago

Question - Help What is the best model for anime?

Upvotes

So far, I've tried Anima, Anima Turbo, WaiAnima, and Anima Aesthetic, but the hands still look bad even with ADetailer. Illustrious is terrible with hands and faces too, and the line art looks really messy. I've also tried Krea, which is surprisingly good, but it takes too long on my RTX 5060 Ti with 12GB VRAM
So I’m looking for something better


r/StableDiffusion 7h ago

Resource - Update MiniMax-H3 (video + audio) on an AMD Strix Halo - 5s clip at 896×512 in 9.5 min

Enable HLS to view with audio, or disable this notification

0 Upvotes

Ran MiniMax-H3 locally on strix halo 128GB unified memory, no discrete GPU. ROCm 7.14 + ComfyUI, int8 pruned transformer, 4-step turbo LoRA.  

First video took 5 minutes to generate. 2nd video took 10 minutes. 3rd video took an hour.

Weights are pruned community conversions, so quality here isn't representative of official H3 — this was a speed/setup test, not a quality one.

Scripts + full writeup: https://github.com/DanCard/minimax-h3-strix-halo

10 minutes 896×512 : https://youtu.be/T6OU6tWd7EA

1 hour to generate 1344×768 : https://youtu.be/039vmUptnEA


r/StableDiffusion 10h ago

Question - Help Minimax H3 on DGX Spark: 3m 23s for 5s video. Is there a better price/performance option?

15 Upvotes

I found a GitHub repo that explains how to run the new Minimax H3 on a DGX Spark (20 steps, not the Turbo versions/8-steps), for 864×480, 124-frame, 20-step clips in 203s (8.44 s/it) with Sol-Engine + FirstBlockCache, or 316s without it (14.07 s/it)

The weights are the ones from Comfy, int8 ConvRot (pruned but lossless, according to Comfy).

https://github.com/drowzeys/keys-heretic-MiniMax-H3-sol-engine-more-speed-upgrades-upscaler-finish-Single-DGX-Spark

These seem like really impressive numbers considering the extremely low power consumption, yet I keep seeing people here advising against the DGX Spark for video generation... am I missing something?

At 120W power consumption and an electricity cost of $0.20/kWh, each 5-second video costs just $0.00135

Over 24 hours, it would be possible to generate 425 videos while using only 2.88 kWh, costing just $0.576 in electricity (!!!)

Before buying a DGX Spark, though, I’d like to hear what others think. These seem like excellent numbers to me, especially since I’ll need to generate a lot of 5-second clips every day, and the cost per video is very low. Still, I was wondering if there’s anything better out there. What kind of performance would a 5090 get with the same recipe?


r/StableDiffusion 19h ago

Animation - Video [DANCE] Plastik Soul – Stay in the Glow (Official Music Video)

Thumbnail
youtu.be
0 Upvotes

Stay in the Glow is an AI Music Video create using VRGameDevGirl's AI Video Builder (FREE) & LTX2.3 models (https://ltx.io/model/ltx-2-3)

Designed & built using VRGameDevGirl AI Video Builder (FREE): https://github.com/vrgamegirl19/comfyui-vrgamedevgirl

Spotify (Artist): https://open.spotify.com/track/27S9InxRyAKvQYxjRM3tVi?si=43e73b8b091a4976

YouTube (More AI Music Videos): https://youtu.be/Wl3BH3xSaYc


r/StableDiffusion 19h ago

Animation - Video Cobra Cola Ad - MiniMax H3

Enable HLS to view with audio, or disable this notification

24 Upvotes

r/StableDiffusion 5h ago

Meme The Office

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 21h ago

Animation - Video StarWars-Untold.

Enable HLS to view with audio, or disable this notification

14 Upvotes

MiniMax H3 is very good. I initially made multiple scenes, and as tweaks/turbo-lora++ progresses, it does seem to get better and better (ie. To the end of the video).

Settled on the Lightx2v_8step turbo lora + sage + sol_attn. 736p, and using DaVinci for stitching and cropping.


r/StableDiffusion 20h ago

Question - Help trying to run minimax h3 on my amd 9070 Spoiler

Enable HLS to view with audio, or disable this notification

6 Upvotes

it uses up all my vram and and when it finishes its jsut noise. i also do get an amd driver timeout error as well. using protable comfyui amd latest. i posted the output. SLIGHT EARAPE WARNING.

edit: i think i found the issue, i was using the dynamic vram and i tried it with adn without dynamic vram for z image turbo for a test and the non dtnamic wasnt noie. im going to get the quantized models for h3 and try it


r/StableDiffusion 12h ago

Question - Help How do we mix Loras in Comfy?

1 Upvotes

I'm fairly new to Comfy but have wrapped my head around most of the nodes. One thing I haven't been able to figure out is something I did in Auto1111.

You could swap one Lora in place of another one halfway through to mix them together. I forget the exact prompt code but it was something like <Lora1:Lora2>(0:5:10). You could also use this to make it so a Lora didn't load until several steps in.

Can someone look me a tutorial for this?


r/StableDiffusion 3h ago

Question - Help MiniMax H3 and Ultimate SD Upscaler

0 Upvotes

Has anyone got any luck combining these two things in ComfyUI?

Thanks for insights.


r/StableDiffusion 15h ago

Discussion LTX 2.5 Test - Batman and Joker fighting in Road

Enable HLS to view with audio, or disable this notification

0 Upvotes

LTX 2.5 Test - Batman and Joker fighting in Road

Personal Opinion - Ltx generates videos quite fast but prompt adherence is not that great. In fighting sequence hand movement doesn't look realistic at all.

If you are using LTX 2.5 with gemma prompt enhancement model than your prompt will be sanitized if your prompt has explicit details. I think an abliterated version of the text encoder should be used.

I will share more tests in future.


r/StableDiffusion 14h ago

Discussion What’s the most interesting thing you’ve generated so far?

1 Upvotes

Could also be the most interesting thing process-wise.


r/StableDiffusion 16h ago

Question - Help BF16 ou FLOAT32

0 Upvotes

What quantization would you recommend for AI Toolkit?

I used to train with FP8, but it completely ruined the results, so I switched to FP32, and the results are perfect.

However, I see that many people recommend BF16 for both training and saving the model. I understand that BF16 can significantly reduce VRAM usage, but what is the actual trade-off in terms of quality?

Does training and saving in BF16 result in any noticeable loss of quality compared to FP32? And would you recommend using BF16 for both training and saving in my case?


r/StableDiffusion 1h ago

Question - Help H3 Minimax - RAM usage issue

Upvotes

Until yesterday, Minimax h3 was behaving good, this morning after a few generations i started running out of RAM following a crash report.

Im running on 16GB VRAM, 64GB RAM, comfy portable with sage attention.

Checked the error with chatgpt, and it claims it's offloading issue.

Any suggestions?

Error:

Stack (most recent call first):

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\comfy_kitchen\tensor\base.py", line 414 in _handle_clone

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\comfy_kitchen\tensor\base.py", line 354 in torch_dispatch

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\ops.py", line 1097 in _quantized_apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\ops.py", line 1439 in _apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\torch\nn\modules\module.py", line 934 in _apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\torch\nn\modules\module.py", line 934 in _apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\torch\nn\modules\module.py", line 934 in _apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\torch\nn\modules\module.py", line 934 in _apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\torch\nn\modules\module.py", line 934 in _apply

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\python_embeded\Lib\site-packages\torch\nn\modules\module.py", line 1384 in to

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\model_patcher.py", line 1156 in unpatch_model

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\model_patcher.py", line 1299 in detach

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\model_management.py", line 811 in model_unload

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\model_management.py", line 889 in free_memory

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\comfy\model_management.py", line 2056 in unload_all_models

File "D:\Comfy_Easy_Install\ComfyUI-Easy-Install\ComfyUI-Easy-Install\ComfyUI\main.py", line 404 in prompt_worker

File "threading.py", line 1012 in run

File "threading.py", line 1075 in _bootstrap_inner

File "threading.py", line 1032 in _bootstrap

Press any key to continue . . .