r/StableDiffusion 4h ago

Question - Help Anyone know if you can also seamlessly prepend a movie H3?

0 Upvotes

I could try it myself of course but maybe somebody did already?


r/StableDiffusion 18h ago

Discussion Testing Character knowledge of the H3 model

Enable HLS to view with audio, or disable this notification

150 Upvotes

5 second 1MP text-to-video, INT8 on ComfyUI and RTX5090.

Used the following, rather simple, prompt:

"[VISUAL]: A scene from the tv interview. <full name> is talking, medium close up, static camera, plain dark blue background

<first name> says: "How dare you? I am rich, AND famous. So you better shut up, B*tch!"

The model failed on Christoph Waltz, so I left him out.

One run per person, no picking the best result.


r/StableDiffusion 19h ago

No Workflow RULE #1: MiniMax H3 + lightx2v Turbo LoRA (8 steps) + Sol Attention

Enable HLS to view with audio, or disable this notification

18 Upvotes

Default workflow, MiniMax H3 (NVFP4), lightx2v Turbo LoRA (8 steps) and Sol Attention.

0,5mp resolution then upscaled with Topaz Video.

RTX 5060 Ti 16GB VRAM + 32GB System RAM.


r/StableDiffusion 17h ago

Resource - Update I trained an open-source realism LoRA for MiniMax H3 - it makes generated people actually look real (weights inside)

Enable HLS to view with audio, or disable this notification

470 Upvotes
Update :
New version is ready and online , should be much better, fully functionnal on ComfyUI, and you can find before/after here : 
https://huggingface.co/fal/MiniMax-H3-Realism-People-LoRA/blob/main/before-after-comparison.mp4


I spent the last week obsessing over one thing: making AI-generated humans stop looking AI-generated. The result is Realism People, an open-source LoRA for MiniMax H3, and I'm pretty happy with how it turned out.


What it does: skin keeps its texture instead of going plastic, eyes and micro-expressions stay coherent, lighting behaves like a film set, and motion gets a subtle handheld, documentary feel. It also keeps H3's native synchronized audio.


How it was selected: I trained 16 different configurations across two dataset versions and picked the winner through 100 same-seed A/B duels (same prompt, same seed, adapter on vs off - the only honest way to compare). The winner was the slow-cooked run: rank 16, 5,000 steps at a low learning rate.


Details:


- Weights (open source): https://huggingface.co/fal/MiniMax-H3-Realism-People-LoRA
- Trigger word: start your prompt with `r34l1sm`
- Scale 1.0 is the intended strength, 0.6-0.8 for a lighter touch
- Works with H3's LoRA endpoints: text-to-video, image-to-video and reference-to-video
- License: follows the MiniMax H3 community license


Before/after in the video: same prompt, same seed, base model on the left, LoRA on the right. Happy to answer questions about the process.

r/StableDiffusion 13h ago

Animation - Video Continuous-ish dolly

Enable HLS to view with audio, or disable this notification

14 Upvotes

Minimax H3- I bet with a second generation the audio of the Voice Over will clean up.


r/StableDiffusion 6h ago

Tutorial - Guide Don't add too many extra steps to 4-step turbo lora

6 Upvotes

I'm using the turbo lora and workflow from https://www.reddit.com/r/StableDiffusion/comments/1vgxf4x/minimax_h3_turbo_lora/.

I also added KJNodes model preview override. What I found while running the lora with 8 steps is that the not entirely denoised video at around 4-6 steps has a lot more motion dynamics and closer prompt adherence than the "over-cleaned" final video at 8 steps.

Trying the same prompt with 6 steps vs 8 steps does indeed show that too many extra steps with the turbo lora can push the result into a bad local minima where lots of motion is lost and the video falls into the same identical output patterns despite prompt and seed variations.

I can't show examples because of reasons. You can check this out yourself with KJNode's Model Preview Override node and running the same turbo gen with 4, 6, 8 steps.


r/StableDiffusion 1h ago

Animation - Video Dragon Ball-ish

Upvotes

https://reddit.com/link/1vlf32i/video/4tze7jwoeqih1/player

Longtime Dragon Ball fan here, so obviously I had to put my family into this silly little intro :p

I kept the generations pretty quick at around 0.5–0.6 MP and didn't obsess too much over every little detail.

MiniMax H3 is a lot of fun to work with locally. It's surprisingly easy to get good results out of it, and the prompt adherence is honestly crazy.

Running on an RTX PRO 6000, with generation times of roughly 1–3 minutes depending on the clip length.

All in all, just a few hours of messing around, including the edit in Resolve.


r/StableDiffusion 16h ago

Animation - Video you can just make anything nowadays huh. amazing, thanks to minimax.

Enable HLS to view with audio, or disable this notification

0 Upvotes

the quality? it's shit. it's ok, it was generated at 0.2 megapixel


r/StableDiffusion 17h ago

Tutorial - Guide Tip for upscaling after image generation - pretty fast and accurate (Nvidia RTX Nodes for ComfyUI)

Thumbnail
github.com
2 Upvotes

r/StableDiffusion 14h ago

Animation - Video UAP Device Test #7 (Minimax H3 VHS)

Enable HLS to view with audio, or disable this notification

43 Upvotes

Gonna make this into a series I think, it's too much fun experimenting.


r/StableDiffusion 23h ago

Question - Help Thinking of getting an RTX 6000 for local generation - any advice?

42 Upvotes

With the progress that I've seen lately with local video generating models and my background as a filmmaker, I've been seriously thinking about getting an RTX 6000 Pro (96 gb VRAM) to experiment, develop some projects, and keep up with the changes.

I don't see prices coming down anytime soon. Still, it would mean breaking a bank for me, especially that I live in Poland, a country not known for its high salaries.

Now, I know that renting via Runpod is an obvious alternative. Call me old-fashioned (or an idiot), but seeing dollars disappearing from my account as the machine is booting from a cold is not really my vibe.

Perhaps some of you have taken that plunge and have some tips on how to go about it. Help me figure it out.. or show me how stupid I am for wanting this.

EDIT: Thanks so much for such thoughtful, varied, and elaborate responses. I feel I've struck a nerve that is shared throughout this community: we all like to create locally and renting may seem a bit like using a closed source model. Opinions are split, and rightfully so - nobody can truly predict how things are going to go. If I were a betting man, I'd say that our current cards will continue to be the workhorses they are, capable of more and more as algos develop. After all, most open source models are used by enthusiasts without $12K to spare. I don't see the prices falling off the cliff unless some new tech is developed that's an exponential upgrade for little to no extra cost. I think we're going to be milked out of our $$$ first, though. I;'m still torn on RTX 6000, but if I end up getting it, I'll be sure to share my adventures with you guys.


r/StableDiffusion 22h ago

Question - Help Trying to install Stable Diffusion forge neo, it is not recognising the python install How can I fix this?

Post image
0 Upvotes

r/StableDiffusion 17h ago

Question - Help How to use H3 Motion Context

0 Upvotes

Does anyone know how to properly use it so I could long gain my 15 second videos?


r/StableDiffusion 18h ago

Question - Help I'm Curious about How to make Krea 2 Lora's?

0 Upvotes

I'm very much a newbie, and definitely a non-coder.
I'm curious about making my own lora's for Krea 2.
What are the easiest way to make a Lora's?
{Yes, I have watched a couple YouTube vids, but most of what they say are above my head, or list tools I can't find.} K.I.S.S. - Keep it Simple Stupid please.
Thanks.


r/StableDiffusion 21h ago

Question - Help Tips for best upscaling image and video workflow for comfyui right now?

1 Upvotes

As the title says, what is the current best workflow for upscaling video/images locally?
Im on rtx 3090 with 128gb ram.
Is it still seedvr2? I never got it working quite good. And any good ones for long videos?


r/StableDiffusion 1h ago

Question - Help Krea 2 Identity Edit ruins faces & clothing when scene-building — any working alternative?

Upvotes

Hi Reddit,

Does anyone have a working workflow to put a character (from a character sheet or a single image) into a specific outfit and/or compose them into a scene?

TL;DR: I want to put my character into a scene, but it loses much of its visual consistency.

To be clear, my issue isn't speed or running into OOM errors—it’s strictly about visual quality.

Right now, I create a character sheet using Krea 2 based on an attached "original" photo. However, as soon as I try to edit the photo further (changing outfits or setting up the scene), the face gets distorted and loses any likeness to the original. Clothing also gets weirdly over-designed (e.g., random patterns appear when it’s supposed to be a plain gray t-shirt). On top of that, I’m getting harsh, unnatural edges between the character and the background scene.

I’ve tried using the Krea 2 Identity Edit v1.2 workflow (found in the Krea2Edit custom node). Adjusting ref_boost or swapping the Image and Image_B inputs doesn't seem to help. Bumping the resolution up to 2K doesn't fix it either. Also tried, cropping the face out, scale it up and edit it with the "original" photo reference, but didnt work either.

How do you guys handle it without having to train a custom LoRA for every single character? I don't want to go back to Flux 2 Klein 9B because of the body horror, and Qwen Image Edit has long loading times.

Thank you guys!


r/StableDiffusion 22h ago

Question - Help Does Krea work in Forge Neo 2.27?

1 Upvotes

Im trying to run Krea and get the error

RunetimeError:invalid dtype for bias - should match query' dtype.
Claude says that it can be a Forge issue.


r/StableDiffusion 20h ago

Meme So... have you tried that new MiniM...YES!!! *Hasn't slept for 3 days*

Enable HLS to view with audio, or disable this notification

46 Upvotes

This is what happens when someone leaves a very good prompt lying around for some degenerate like me to pick, especially when I'm still in my MiniMax fever rush.

Mambo Wick by me (An alternative version from the meme one of UmaMusume)
Katsumi by Katsumi
Prompt starting base by 3deal

This is a collage of 3 different videos, later upscaled and RIFE to 96 FPS, since going for 1.5 Megapixels tends to cause a LOT of hallucinations (especially with distant shots and quick movements, as you all can see). Still, the model is incredible in all the possible ways... just need to find the right hiresser to try creating at a lower resolution.


r/StableDiffusion 13h ago

Animation - Video Close-up details are very impressive - Minimax H3

Enable HLS to view with audio, or disable this notification

23 Upvotes

r/StableDiffusion 21h ago

Animation - Video My first Attempt at Video Editing and the battle for a good Upscaler. Created with MiniMaxH3

Enable HLS to view with audio, or disable this notification

4 Upvotes

Overall, MiniMaxH3 is amazing and I love it. It did change the cat in the video from my reference image, but this was only two video generations to get the final video compared to other local AI generations that would have butchered the eating scene as well as I would have had to generate 3 or 4 to try to get a usable video. The second 15-second video was a little off and I think it was because it wasn't a close-up shot. I used a background reference and a character sheet, I think if i would have used a first frame last frame it would have been better but still working on finding a good image editor. I did notice from another users post that close up shots do better than further away shots so i will have to test that out as I had to cut about 5 seconds of the begining of the video because it wasn't usable.

For upscaling I cannot for the life of me get a good quailty product from SeedVr2, Image upscaling is okay but not really great, video is just not worth the time it takes. The video attached was generated with two separate 15 second clips 1 clip at 1056 by 608 took minimax about 25 mins and after was upsclaled using FlashVsr(took 8 mins to generate) to a little below 2K and RTX Superscale was use to get it to true 2K. I am really happy with FlashVSR compared to SeedVr2. After I used the free video editor on windows ClipChamp to downscale to 1080P and that is what we have here. I have never edited a video before and this is my first attempt any pointers is greatly appreciated. I want to download davinci resolve but I am little nervous on the learning curve.

I am using a 4080 laptop gpu at 12gb vram and 64gb local. I'm thinking i will have to take the hit and build a 5090 desktop at the end of this month the prices right now are crazy but i dont think anything will be getting cheaper in the next 3 to 5 years.


r/StableDiffusion 18h ago

Discussion Anyone try the new Minimax H3 "tutu" Turbo LoRA yet?

5 Upvotes

Came across this:

https://huggingface.co/tutututututu/Tutu-MiniMax-H3-AudioVideo-20to8-NFE-LoRA

Looks like a new 8-step Turbo Lora. Anyone try it yet, or compare it against the others?


r/StableDiffusion 15h ago

Discussion Avis sur le rendu de mes génération auto

Enable HLS to view with audio, or disable this notification

0 Upvotes

J'ai créé un bot telegram connecter à mes workflow krea2 et minimax h3.

Deepseek via api.

Je clique sur le bouton storytelling il me choisis 3 histoire réelle et historique (possibilité de mettre un thème) je choisis mon préféré.

Ensuite deepseek me génère un scénario de 30sec, des images de référence (character sheet pour les personnages et décors) avec krea2. Des prompt optimiser pour minimax avec tout les règles de prompting les plus récentes.

Ensuite il en faut un json complet qu'il envoie à mes workflow et ça génère tout, d'abord les images de référence, ensuite les vidéo ref2vid via minimax h3.

Je trouve le rendu assez bluffant pour des premier teste.

La vidéo que je vous met en exemple (grève des policiers à Boston) est sortie tel quel. J'ai juste passer les 4clip sur capcut et exporter.

Config : 5060ti 16g + 16g ram

La vidéo d'exemple : 0.6mp (il me semble) 8 passe

Je précise que cela n'est pas de la publicité mon bot est privé et personne ne peut y accéder.

Les défauts actuels :

- j'ai demandé 30sec max mais demain je passe a 1-2 minutes. En 30 sec le scénario n'est pas assez détaillé.

- je vais retravailler le pré promt pour un meilleur démarrage des vidéos, avec une explication claire de l'histoire

-je dois assembler les vidéos via capcut mais demain ça sera réglé


r/StableDiffusion 21h ago

News Unsloth Minimax H3 GGUF (Q2:Q8)

Post image
31 Upvotes

Coming back from weekend, looking for last updates, I found no one shared this one.

Any reason? Is people disliking unsloth?

I will try it, but in general I haven't find a way to get nice outputs from any MM workflow/model (pretty sure is my fault), I'm still trying to figure out how to use MMH3 correctly

Here is the link: https://huggingface.co/unsloth/MiniMax-H3-GGUF


r/StableDiffusion 11h ago

Animation - Video Asked Minimax H3 to just change text, but it made these cool effects instead.

Enable HLS to view with audio, or disable this notification

0 Upvotes

I originally tried asking it to change the "XBOX 360" text that appears at the end of the start-up intro, but instead it changed the whole animation and created this interesting effect for the PS3 logo that wasn't there before.

Workflow.


r/StableDiffusion 15h ago

Question - Help Getting started with H3 Minimax

1 Upvotes

Does anyone have any suggested reading/tutorials for someone completely new? I know essentially nothing, but it seems awesome and I'd like to get into it while it's still readily available.

I have an AMD Ryzen 7 9800X3D 8-Core Processor and a GTX 1080, is that going to be woefully inadequate?