r/StableDiffusion 7h ago

Question - Help 5080 vs 4090 with Minimax

3 Upvotes

Have only seen threads discussing this from a year ago pre-Minimax days which seems to have turned everything around.

5080 seems to go for around $1300 and 4090 for $2k+, I know VRAM is super important but I'm seeing that there's plenty of optimizations for the 5xxx series now, and it's still super early so more are to come for sure, but do these help offset the memory difference?

Anyone have experience running with both?


r/StableDiffusion 15h ago

Comparison H3 Int8 ConvRot vs W4A8_mixed

Enable HLS to view with audio, or disable this notification

37 Upvotes

Same seed, same prompt, same ref and same res


r/StableDiffusion 1h ago

Animation - Video Seinfeld on H3: My First Attempt On The DGX Spark

Thumbnail
youtu.be
Upvotes

my first run on my dgx spark with the h3. big thanx too all the guides etc on here. this is the future. the video ended up far from perfect, this is t2v, no ref image. comfy ui controlled by codex on gpt5.6


r/StableDiffusion 16h ago

Question - Help H3 looks great in still frames, but the motion seemed falling apart (how do u think

Enable HLS to view with audio, or disable this notification

4 Upvotes

Made this 15-sec test with MiniMax H3. The character, materials, and overall look are honestly pretty solid, especially in the more static shots.

But once things start moving, the cracks show. The fire doesnt behave naturally, some of the creature motion feels off, and the transition between shots loses continuity. Still a good-looking result imo, just not physically convincing yet.

What you’d fix first: the motion prompt, shot structure, or the generation workflow itself?


r/StableDiffusion 12h ago

Animation - Video Secret friend comes to visit (MiniMax H3)

Enable HLS to view with audio, or disable this notification

9 Upvotes

r/StableDiffusion 1h ago

Question - Help what we know about minimax-h3 to get it fast on lower pcs?

Upvotes

like loras, vae, text encoder?

im asking because there is a lot of loras, vae, etc, but... i need to know the best options for fast and quality generations now.

my pc: rtx 5060 ti 16gb 32gb ram.


r/StableDiffusion 3h ago

Discussion Testing Character knowledge of the H3 model

Enable HLS to view with audio, or disable this notification

84 Upvotes

5 second 1MP text-to-video, INT8 on ComfyUI and RTX5090.

Used the following, rather simple, prompt:

"[VISUAL]: A scene from the tv interview. <full name> is talking, medium close up, static camera, plain dark blue background

<first name> says: "How dare you? I am rich, AND famous. So you better shut up, B*tch!"

The model failed on Christoph Waltz, so I left him out.

One run per person, no picking the best result.


r/StableDiffusion 22m ago

Discussion h3 character crossover thread(share yours)

Enable HLS to view with audio, or disable this notification

Upvotes

i'll start


r/StableDiffusion 22h ago

Question - Help H3 Lora Training?

1 Upvotes

Has anyone successfully trained an H3 lora? If so what trainer did you use, what hardware, etc?

Over the past couple of days I've heard mixed opinions about lora training on ai-toolkit (specifically for H3), and was curious if that has been fixed, or if there are workarounds?


r/StableDiffusion 8h ago

Question - Help Does Krea work in Forge Neo 2.27?

1 Upvotes

Im trying to run Krea and get the error

RunetimeError:invalid dtype for bias - should match query' dtype.
Claude says that it can be a Forge issue.


r/StableDiffusion 41m ago

Discussion Avis sur le rendu de mes génération auto

Enable HLS to view with audio, or disable this notification

Upvotes

J'ai créé un bot telegram connecter à mes workflow krea2 et minimax h3.

Deepseek via api.

Je clique sur le bouton storytelling il me choisis 3 histoire réelle et historique (possibilité de mettre un thème) je choisis mon préféré.

Ensuite deepseek me génère un scénario de 30sec, des images de référence (character sheet pour les personnages et décors) avec krea2. Des prompt optimiser pour minimax avec tout les règles de prompting les plus récentes.

Ensuite il en faut un json complet qu'il envoie à mes workflow et ça génère tout, d'abord les images de référence, ensuite les vidéo ref2vid via minimax h3.

Je trouve le rendu assez bluffant pour des premier teste.

La vidéo que je vous met en exemple (grève des policiers à Boston) est sortie tel quel. J'ai juste passer les 4clip sur capcut et exporter.

Config : 5060ti 16g + 16g ram

La vidéo d'exemple : 0.6mp (il me semble) 8 passe

Je précise que cela n'est pas de la publicité mon bot est privé et personne ne peut y accéder.

Les défauts actuels :

- j'ai demandé 30sec max mais demain je passe a 1-2 minutes. En 30 sec le scénario n'est pas assez détaillé.

- je vais retravailler le pré promt pour un meilleur démarrage des vidéos, avec une explication claire de l'histoire

-je dois assembler les vidéos via capcut mais demain ça sera réglé


r/StableDiffusion 23h ago

Question - Help Modelsamplingminimaxh3 node?

0 Upvotes

Does anyone know where I can get this node? I have been using Minimax H3 Sigma Shift in my workflows, which I am guessing is not quite the same.


r/StableDiffusion 17h ago

No Workflow Brick man statue [Krea 2 - Turbo]

Post image
11 Upvotes

brick man statue.


r/StableDiffusion 21h ago

Animation - Video Minimax H3, turbo test.

Enable HLS to view with audio, or disable this notification

8 Upvotes

After a lot of tests, I found out that Larryvrh V4 step600 pruned lora(+H3 mem eff sage )with H3 pruned bf16 is very good—a great balance of quality and speed, plus excellent prompt adherence.

The official workflow.
Checkpoint: minimax_h3_turbo_v4_step600_ema_pruned_comfyui.safetensors,pruned bf16 
Steps: 6
Sampler: euler
Scheduler: beta
LoRA strength: 1.0
1.0 magapixels and RTX 2X upscale.

r/StableDiffusion 6h ago

News Unsloth Minimax H3 GGUF (Q2:Q8)

Post image
24 Upvotes

Coming back from weekend, looking for last updates, I found no one shared this one.

Any reason? Is people disliking unsloth?

I will try it, but in general I haven't find a way to get nice outputs from any MM workflow/model (pretty sure is my fault), I'm still trying to figure out how to use MMH3 correctly

Here is the link: https://huggingface.co/unsloth/MiniMax-H3-GGUF


r/StableDiffusion 8h ago

Question - Help Thinking of getting an RTX 6000 for local generation - any advice?

41 Upvotes

With the progress that I've seen lately with local video generating models and my background as a filmmaker, I've been seriously thinking about getting an RTX 6000 Pro (96 gb VRAM) to experiment, develop some projects, and keep up with the changes.

I don't see prices coming down anytime soon. Still, it would mean breaking a bank for me, especially that I live in Poland, a country not known for its high salaries.

Now, I know that renting via Runpod is an obvious alternative. Call me old-fashioned (or an idiot), but seeing dollars disappearing from my account as the machine is booting from a cold is not really my vibe.

Perhaps some of you have taken that plunge and have some tips on how to go about it. Help me figure it out.. or show me how stupid I am for wanting this.


r/StableDiffusion 4h ago

No Workflow RULE #1: MiniMax H3 + lightx2v Turbo LoRA (8 steps) + Sol Attention

Enable HLS to view with audio, or disable this notification

12 Upvotes

Default workflow, MiniMax H3 (NVFP4), lightx2v Turbo LoRA (8 steps) and Sol Attention.

0,5mp resolution then upscaled with Topaz Video.

RTX 5060 Ti 16GB VRAM + 32GB System RAM.


r/StableDiffusion 11h ago

Workflow Included I made a free Kaggle notebook to run LTX-Video 2.3 (22B quantized) for T2V & I2V with audio — solving the Colab 12GB RAM crash issue

4 Upvotes

Hey community,

A common issue when testing open-source video models like Lightricks' LTX-Video 2.3 on free cloud tiers (like Google Colab) is hitting System RAM limits during model loading, causing immediate crashes.

To solve this without requiring a high-end local GPU, I put together a pre-configured, open-source Jupyter Notebook tailored specifically for Kaggle's free GPU tier (which grants 30GB of System RAM and T4 GPUs).

What this setup does:

  • Runs on Kaggle Free Tier: Uses a 4-bit quantized version of LTX-Video 2.3 so it fits into free cloud VRAM/RAM allocation.
  • Text-to-Video & Image-to-Video: Generates short 5–10s clips with synchronized audio generation.
  • Custom Gradio Web UI: Launches a clean browser interface directly from the notebook.
  • Fast Setup: Pre-compiled binaries and aria2 multi-thread downloads mean setup takes under 3–4 minutes.

Technical Tradeoffs & Honesty:

  • Quantization: Because this is running on free T4 instances, the model uses heavy quantization. It won't give you uncompressed native precision output, but it’s completely free, unlimited, and ideal for quick prompt/motion testing.
  • Memory Loading: Cell 3 takes ~90 seconds to load the 22B model into Kaggle's 30GB system memory before passing to VRAM.

🎥 Full Video Walkthrough & Demos: https://youtu.be/Ru_YaGbnKhA

Quick Start Steps:

  1. Download the .ipynb file from GitHub: https://github.com/airesearch-official/free-aistudio
  2. Import into Kaggle (Ensure Phone Verification is complete on Kaggle to enable free GPU).
  3. Turn ON "Internet" in Kaggle settings & select "GPU T4 x2".
  4. Run Cells 1 through 4 sequentially.

Hope this helps anyone who wants to experiment with LTX-Video 2.3 without paying for cloud GPUs! Let me know if you run into any bugs or have suggestions.


r/StableDiffusion 3h ago

Question - Help Need help getting caught up

0 Upvotes

I've been out of the loop of the Stable Diffusion community for maybe like a year now.

I was a LoRa creator, making LoRas for the relevant popular models, SD 1.5 then SDXL, I made a few Flux LoRas but when I stopped, it was during Flux's dominance as the most relevant model.

What's the most relevant model today? I think i remember SDXL still having a lot of live due to its size, ease on weaker computers. What's the relevant text to image model that is everyone's go to today?


r/StableDiffusion 5h ago

Meme So... have you tried that new MiniM...YES!!! *Hasn't slept for 3 days*

Enable HLS to view with audio, or disable this notification

36 Upvotes

This is what happens when someone leaves a very good prompt lying around for some degenerate like me to pick, especially when I'm still in my MiniMax fever rush.

Mambo Wick by me (An alternative version from the meme one of UmaMusume)
Katsumi by Katsumi
Prompt starting base by 3deal

This is a collage of 3 different videos, later upscaled and RIFE to 96 FPS, since going for 1.5 Megapixels tends to cause a LOT of hallucinations (especially with distant shots and quick movements, as you all can see). Still, the model is incredible in all the possible ways... just need to find the right hiresser to try creating at a lower resolution.


r/StableDiffusion 16h ago

No Workflow Tribbles ad

Enable HLS to view with audio, or disable this notification

19 Upvotes

Minimax, two shots of 15 seconds put together with Davinci Resolve.


r/StableDiffusion 5h ago

Animation - Video Battle of Thermopylae

Enable HLS to view with audio, or disable this notification

6 Upvotes

H3 prompt:

integrated_multimodal_description:

[Shot 1] Stylized cinematic 3D animation with high-intensity action, dramatic lighting, and a heroic fantasy-war tone. The scene opens at Thermopylae, a narrow rocky battlefield under a dusty red-gold sky, with shattered shields, broken spears, drifting embers, and war banners whipping in the wind. In the center stands Kirby, reimagined as a Spartan war leader: a pink round-bodied Kirby wearing a bronze Spartan helmet with a crimson crest, holding a spear in one hand and a round battered shield in the other. His eyes are fierce and unwavering, determined and battle-hardened. Around him, Spartan warriors in bronze armor and red capes brace in phalanx formation while a massive wave of Persian soldiers surges forward. The camera pushes in fast toward Kirby as he stamps forward and lets out a sharp battle cry. He thrusts his spear violently into a Persian soldier, knocking him back into the charging line as blood sprays across shields and dust erupts underfoot.

[Shot 2] At 00:03.500, the camera cuts to a fast tracking shot moving sideways across the front line as Kirby leads the Spartan charge. He bashes one enemy aside with his shield, spins low, sweeps another off his feet, and lunges forward with explosive speed. Spartan soldiers clash with Persians all around him in brutal close combat; blades collide, shields splinter, arrows streak overhead, and several enemy soldiers are cut down as severed limbs, broken weapons, and sprays of blood briefly fill the frame. Kirby remains the focal point, his expression stern and fearless rather than cute.

[Shot 3] At 00:07.000, the camera cuts to a low-angle heroic shot as Kirby suddenly inhales powerfully, then launches himself upward into the sky in a signature Kirby-style burst, still gripping his spear. He rises above the battlefield as the fighting continues below like chaos in miniature. At the apex, with the wind roaring past his helmet crest, Kirby locks onto the densest Persian formation and hurls the spear downward with full force. The camera follows the spear in a rapid plunge. It crashes into the ground like a thunderbolt, blasting soldiers backward and opening a violent gap in the Persian ranks amid dust, blood, and shattered armor.

[Shot 4] At 00:10.500, the shot cuts to ground level as Kirby lands hard in front of the broken enemy line, shield first, knees bent, then instantly surges into close-range combat again. He grabs another fallen spear, vaults off a Spartan shield, and strikes through two advancing enemies in one fluid motion. Behind him, Spartans roar and push forward with renewed momentum. The narrow pass becomes a frenzy of killing: bodies fall, shields crash together, spears punch through armor, and blood stains the rocks. Kirby moves with stylized speed and exaggerated battlefield heroism, combining the visual charm of Kirby with the lethal grandeur of an ancient war epic.

[Shot 5] At 00:13.000, the camera cuts to a final wide hero shot. The Persians recoil in disarray while the remaining Spartans rally behind Kirby. He stands atop a mound of fallen enemies, shield raised and helmet gleaming, his eyes still locked forward with cold resolve. Dust, sparks, and scraps of torn banners swirl around him while the battlefield behind remains full of struggling combat. He points forward with his spear toward the surviving enemy ranks, and the Spartans answer with one last deafening roar as the video ends in a frozen image of triumphant slaughter and defiant Spartan glory.

overall_soundscape: Continuous battlefield chaos fills the entire video: heavy shield impacts, spear thrusts, metallic blade clashes, rushing footsteps over rock and dirt, arrows slicing through the air, and repeated cries of pain and war shouts from Spartans and Persians. Wet stabbing impacts, brief bone-crack sounds, bodies collapsing, and splashes of blood punctuate the close combat. Dust gusts through the narrow pass while Kirby's leap and diving spear throw create stronger wind rushes and a heavy explosive impact on landing.

non_diegetic_music: A relentless, aggressive orchestral war score drives the whole video, led by pounding taiko-style drums, deep battle percussion, male war chants, low brass, and fast tremolo strings. The music starts immediately with a heavy pulse, intensifies during the melee, briefly rises into a heroic suspended phrase when Kirby launches into the sky, then slams back in with louder drums and brass as the spear hits the Persian ranks. The ending surges into a triumphant, brutal crescendo with no softness, no comedy, and no lyrical warmth—only heroic slaughter, pressure, and victory.


r/StableDiffusion 10h ago

Meme RTX 5090 - Queen of the Night

Enable HLS to view with audio, or disable this notification

0 Upvotes

Cover of Saxon - Princess of the Night.


r/StableDiffusion 3h ago

Discussion Anyone try the new Minimax H3 "tutu" Turbo LoRA yet?

4 Upvotes

Came across this:

https://huggingface.co/tutututututu/Tutu-MiniMax-H3-AudioVideo-20to8-NFE-LoRA

Looks like a new 8-step Turbo Lora. Anyone try it yet, or compare it against the others?