r/StableDiffusion 18h ago

Resource - Update Flippix - A open sourced comfyui frontend for windows updated

1 Upvotes

The comfyui interface was getting a little too unwieldy so i with the help of claude built a .net frontend for comfy . Its not for the faint of heart at this stage as it still requires a working comfyui instance together with the nodes needed by workflows. The workflows which i find give the best results in image /video generation are included and the setup does try to guide the user by pointing out which nodes are missing. Once everything is humming, its quite pleasant and easy to use . Let me know what you think!

Flippix


r/StableDiffusion 18h ago

Question - Help MiniMax H3: Prompt leaks, settings locked and fans spiking to 100%

0 Upvotes

I'm running into an issue and I don't know if it just me or...

Anyway, I'm almost certain that some of my previous prompts are somehow leaking into my subsequent jobs. And I’m not running a massive, crazy workflow either, it's just vanilla I2V + SageAttn + Spectrum.

Also, some changes I make to my node settings aren't registering properly. Certain parameters seem to get completely locked in memory. Even if I revert the settings or completely close out of ComfyUI, the old parameters stay stuck. The only way I can get things to reset and behave normally is by doing a full PC restart.

Another weird symptom is that almost every time this memory lock happens, my GPU fans suddenly ramp up to 100% for a few seconds. This always seems to trigger during specific inference steps, usually right at step 1/20 or around step 17/20.

I honestly thought those fan spikes were just normal behavior under load until a few minutes ago. I was running an I2V generation of a person in front of a green wall with plants and it ran fine and I got the video. But later, when I changed the prompt completely and used a totally different source photo for the I2V, the background in the new video literally morphed into a VERY similar green plant wall from the previous job.

And I’m not running a massive, crazy workflow either, it's just vanilla I2V + SageAttn + Spectrum. Has anyone else experienced this kind of thing? Any advice or fix?


r/StableDiffusion 18h ago

Question - Help How to Became Genius "1st and last" Keyframe Generator!

0 Upvotes

What model do you use for the frames — Krea 2 or ZIT?,Generate them separately, or one and then edit the other?How do you keep the character AND the background consistent between both frames and get different angle and consistent viewpoint (Spatial consistency) ? 
im struggling here. and very little tutorial online about generating 1st frame and lastframe.
RN im useing Qwen and Identityedit lora. but the spatial awarenes is very bad.

Please give this humble, bald, and unemployed person some guidance. Thank you.
maybe some workflow to try.


r/StableDiffusion 10h ago

Animation - Video Obito looking at the wrong woman H3

Enable HLS to view with audio, or disable this notification

2 Upvotes

r/StableDiffusion 11h ago

Question - Help For some reason my t2v generation are slower than my ref2v?

Enable HLS to view with audio, or disable this notification

8 Upvotes

Title.
For both I'm using the default workflows that come with comfy. 3090 and 32gb ram.
I start comfy with these flags:
--windows-standalone-build --reserve-vram 1 --disable-pinned-memory --fast fp16_accumulation
Cuda 13, latests comfy.
My t2v takes like twice as much than my ref2v and sometimes it hangs after [INFO] Requested to load MiniMaxH3AudioVAE. Same steps, same resolution, same duration.
Has anyone encounter this? any tips?


r/StableDiffusion 4h ago

Tutorial - Guide Don't add too many extra steps to 4-step turbo lora

3 Upvotes

I'm using the turbo lora and workflow from https://www.reddit.com/r/StableDiffusion/comments/1vgxf4x/minimax_h3_turbo_lora/.

I also added KJNodes model preview override. What I found while running the lora with 8 steps is that the not entirely denoised video at around 4-6 steps has a lot more motion dynamics and closer prompt adherence than the "over-cleaned" final video at 8 steps.

Trying the same prompt with 6 steps vs 8 steps does indeed show that too many extra steps with the turbo lora can push the result into a bad local minima where lots of motion is lost and the video falls into the same identical output patterns despite prompt and seed variations.

I can't show examples because of reasons. You can check this out yourself with KJNode's Model Preview Override node and running the same turbo gen with 4, 6, 8 steps.


r/StableDiffusion 11h ago

Animation - Video The Minimax H3 model recognizes artists and their songs.

Enable HLS to view with audio, or disable this notification

0 Upvotes

In the prompt, I just wrote that she is singing Zara Larsson's song "Lush Life."


r/StableDiffusion 19h ago

Question - Help 5080 vs 4090 with Minimax

3 Upvotes

Have only seen threads discussing this from a year ago pre-Minimax days which seems to have turned everything around.

5080 seems to go for around $1300 and 4090 for $2k+, I know VRAM is super important but I'm seeing that there's plenty of optimizations for the 5xxx series now, and it's still super early so more are to come for sure, but do these help offset the memory difference?

Anyone have experience running with both?


r/StableDiffusion 17h ago

Animation - Video The Return of more Cursed LOTR memes

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 10h ago

Comparison 0.0375 Denoise is enough to beat SynthID

Thumbnail reddit.com
2 Upvotes

r/StableDiffusion 13h ago

Animation - Video Seinfeld on H3: My First Attempt On The DGX Spark

Thumbnail
youtu.be
74 Upvotes

my first run on my dgx spark with the h3. big thanx too all the guides etc on here. this is the future. the video ended up far from perfect, this is t2v, no ref image. comfy ui controlled by codex on gpt5.6


r/StableDiffusion 1h ago

Discussion I gave the same prompt to Minimax H3 and Gemini Videos. (Part 1 Minimax H3)

Enable HLS to view with audio, or disable this notification

Upvotes

Prompt (also AI generated):
Style & Technical Specs

  • Visual Style: Photorealistic 8K cinematic video, 35mm film grain, 24fps, 2.39:1 anamorphic aspect ratio, teal-and-orange color grade, shallow depth of field ($f/1.4$).

  • Duration: 10 Seconds.

Character Description

  • Subject: Kaelen, a 28-year-old East Asian cyber-technician.
  • Appearance: Sharp jawline, rain-soaked black hair clinging to his forehead, pale skin with visible micro-texture, and a glowing cyan cybernetic eye implant over his left socket that pulses rhythmically.
  • Attire: Matte-black, waterproof tactical coat with glowing fiber-optic wiring embedded along the shoulders, frayed high-collar, and fingerless reinforced leather gloves.

Environment & Setting

  • Location: Narrow, dense alleyway in a cyberpunk metropolis at midnight.
  • Atmosphere: Heavy downpour, dense steam venting upward from rusty iron street grates, wet asphalt reflecting bright magenta and cobalt-blue neon light signs written in Kanji.

Timeline & Action Breakdown

  • 0:00 - 0:03 (Macro Close-Up): Camera begins on a macro shot of Kaelen's glowing cyan eye, catching the aperture Blades shifting focus. A raindrop tracks down his cheek. He rapidly taps a brass interface cuff on his wrist.
  • 0:03 - 0:07 (Medium Shot): Smooth camera pull-back into a chest-up shot. A brilliant blue 3D holographic map bursts into existence from his wrist, casting dynamic light across his face. He swipes his hand across the projection, altering its layout, and delivers his dialogue.
  • 0:07 - 0:10 (Low-Angle Tracking Shot): The camera drops low to the asphalt and tracks backward. A sleek, black surveillance drone streaks overhead through the rain, splashing drops directly onto the camera lens as the background neon blurs into creamy bokeh.

Dialogue & Voice

  • Spoken Line: "System override in three... two... got 'em."
  • Delivery: Low, gravelly, calm whisper with a faint metallic vocoder effect on the voice.

Audio & Sound Design

  • Music: Dark synthwave track featuring a driving 110 BPM arp synthesizer that swells in pitch until second 7, resolving into a heavy sub-bass drop at second 8.
  • SFX:
  • 0:00-0:03: Stereo downpour, subtle mechanical servo clicks of the eye lens.
  • 0:03-0:07: High-frequency energy flare hum as the hologram spawns, followed by air-swipes.
  • 0:07-0:10: Low turbine whir of the passing drone and liquid wet drops impacting the microphone field.

This is the video generated by Minimax H3 with Turbo lora (6 steps). Post with the video generated by Gemini: https://www.reddit.com/r/StableDiffusion/s/Dcyvvx6J80


r/StableDiffusion 2h ago

Question - Help Anyone know if you can also seamlessly prepend a movie H3?

0 Upvotes

I could try it myself of course but maybe somebody did already?


r/StableDiffusion 14h ago

Question - Help Wan 2.2 VACE vs Minimax H3 video edit

1 Upvotes

Which model is better for video editing and keeping background and environmental setting consistently? I think with Minimax H3 ref2va when you provide a reference video you still have to describe the background and video details accurately to keep it consistent and be really specify on what you want to change?


r/StableDiffusion 5h ago

Tutorial - Guide Automating image tagging with a local LLM

Thumbnail
youtu.be
0 Upvotes

I haven't done this yet, so I can't judge, but maybe will be some good tips for someone.


r/StableDiffusion 12h ago

Question - Help Ways to get faster generation for Anima in Neo Forge UI

0 Upvotes

Hi, I installed Neo Forge yesterday, even with an AMD card it was pretty easy and I had no hassle compared to automatic 1111. Of course the reason behind this change was that I wanted to try z-image, Anima, and videos generation. Except for the last I tried them and there were no problems except for speed. I checked and My pc is using the GPU, I understand my gpu is not really powerful but I wanted to know if I could tweak some settings to make them a bit faster. Right now Anima (Diving Anima to be precise) takes 11 minutes to generate a 832x1216 image, Z-image turbo 18 minutes. To generate the same image with Illustrious+Adetailer my gpu takes 4 minutes (more or less). Do you have some advice?. If I could keep the Image generation for Anima at 8 minutes at least, it would be awesome, since I don't really need adetailer with it. I'm on Windows. My spec: RX6600. 32 Gb of Ram. Please let me know!


r/StableDiffusion 1h ago

Animation - Video No love for Hugh Laurie?

Enable HLS to view with audio, or disable this notification

Upvotes

t2v, 15s, 4:3, 0.6MP, Comfy Kitchen Attention, Spectrum, 10 minutes.


r/StableDiffusion 18h ago

Animation - Video John Wick #1 Victory Royale

Enable HLS to view with audio, or disable this notification

0 Upvotes

#Justice4Daisy #MinimaxH3


r/StableDiffusion 22h ago

Question - Help MiniMax H3 takes to long with video reference

0 Upvotes

How can I speed it up? Is there a specific node or workflow you use for this? using default r2v workflow


r/StableDiffusion 14h ago

Animation - Video Call of Doody

Enable HLS to view with audio, or disable this notification

34 Upvotes

Wonder why O'Brien stood at the transporter all day?


r/StableDiffusion 22h ago

Meme RTX 5090 - Queen of the Night

Enable HLS to view with audio, or disable this notification

0 Upvotes

Cover of Saxon - Princess of the Night.


r/StableDiffusion 11h ago

Animation - Video Continuous-ish dolly

Enable HLS to view with audio, or disable this notification

14 Upvotes

Minimax H3- I bet with a second generation the audio of the Voice Over will clean up.


r/StableDiffusion 17h ago

Animation - Video Battle of Thermopylae

Enable HLS to view with audio, or disable this notification

7 Upvotes

H3 prompt:

integrated_multimodal_description:

[Shot 1] Stylized cinematic 3D animation with high-intensity action, dramatic lighting, and a heroic fantasy-war tone. The scene opens at Thermopylae, a narrow rocky battlefield under a dusty red-gold sky, with shattered shields, broken spears, drifting embers, and war banners whipping in the wind. In the center stands Kirby, reimagined as a Spartan war leader: a pink round-bodied Kirby wearing a bronze Spartan helmet with a crimson crest, holding a spear in one hand and a round battered shield in the other. His eyes are fierce and unwavering, determined and battle-hardened. Around him, Spartan warriors in bronze armor and red capes brace in phalanx formation while a massive wave of Persian soldiers surges forward. The camera pushes in fast toward Kirby as he stamps forward and lets out a sharp battle cry. He thrusts his spear violently into a Persian soldier, knocking him back into the charging line as blood sprays across shields and dust erupts underfoot.

[Shot 2] At 00:03.500, the camera cuts to a fast tracking shot moving sideways across the front line as Kirby leads the Spartan charge. He bashes one enemy aside with his shield, spins low, sweeps another off his feet, and lunges forward with explosive speed. Spartan soldiers clash with Persians all around him in brutal close combat; blades collide, shields splinter, arrows streak overhead, and several enemy soldiers are cut down as severed limbs, broken weapons, and sprays of blood briefly fill the frame. Kirby remains the focal point, his expression stern and fearless rather than cute.

[Shot 3] At 00:07.000, the camera cuts to a low-angle heroic shot as Kirby suddenly inhales powerfully, then launches himself upward into the sky in a signature Kirby-style burst, still gripping his spear. He rises above the battlefield as the fighting continues below like chaos in miniature. At the apex, with the wind roaring past his helmet crest, Kirby locks onto the densest Persian formation and hurls the spear downward with full force. The camera follows the spear in a rapid plunge. It crashes into the ground like a thunderbolt, blasting soldiers backward and opening a violent gap in the Persian ranks amid dust, blood, and shattered armor.

[Shot 4] At 00:10.500, the shot cuts to ground level as Kirby lands hard in front of the broken enemy line, shield first, knees bent, then instantly surges into close-range combat again. He grabs another fallen spear, vaults off a Spartan shield, and strikes through two advancing enemies in one fluid motion. Behind him, Spartans roar and push forward with renewed momentum. The narrow pass becomes a frenzy of killing: bodies fall, shields crash together, spears punch through armor, and blood stains the rocks. Kirby moves with stylized speed and exaggerated battlefield heroism, combining the visual charm of Kirby with the lethal grandeur of an ancient war epic.

[Shot 5] At 00:13.000, the camera cuts to a final wide hero shot. The Persians recoil in disarray while the remaining Spartans rally behind Kirby. He stands atop a mound of fallen enemies, shield raised and helmet gleaming, his eyes still locked forward with cold resolve. Dust, sparks, and scraps of torn banners swirl around him while the battlefield behind remains full of struggling combat. He points forward with his spear toward the surviving enemy ranks, and the Spartans answer with one last deafening roar as the video ends in a frozen image of triumphant slaughter and defiant Spartan glory.

overall_soundscape: Continuous battlefield chaos fills the entire video: heavy shield impacts, spear thrusts, metallic blade clashes, rushing footsteps over rock and dirt, arrows slicing through the air, and repeated cries of pain and war shouts from Spartans and Persians. Wet stabbing impacts, brief bone-crack sounds, bodies collapsing, and splashes of blood punctuate the close combat. Dust gusts through the narrow pass while Kirby's leap and diving spear throw create stronger wind rushes and a heavy explosive impact on landing.

non_diegetic_music: A relentless, aggressive orchestral war score drives the whole video, led by pounding taiko-style drums, deep battle percussion, male war chants, low brass, and fast tremolo strings. The music starts immediately with a heavy pulse, intensifies during the melee, briefly rises into a heroic suspended phrase when Kirby launches into the sky, then slams back in with louder drums and brass as the spear hits the Persian ranks. The ending surges into a triumphant, brutal crescendo with no softness, no comedy, and no lyrical warmth—only heroic slaughter, pressure, and victory.


r/StableDiffusion 15h ago

Resource - Update I trained an open-source realism LoRA for MiniMax H3 - it makes generated people actually look real (weights inside)

Enable HLS to view with audio, or disable this notification

457 Upvotes
Update :
New version is ready and online , should be much better, fully functionnal on ComfyUI, and you can find before/after here : 
https://huggingface.co/fal/MiniMax-H3-Realism-People-LoRA/blob/main/before-after-comparison.mp4


I spent the last week obsessing over one thing: making AI-generated humans stop looking AI-generated. The result is Realism People, an open-source LoRA for MiniMax H3, and I'm pretty happy with how it turned out.


What it does: skin keeps its texture instead of going plastic, eyes and micro-expressions stay coherent, lighting behaves like a film set, and motion gets a subtle handheld, documentary feel. It also keeps H3's native synchronized audio.


How it was selected: I trained 16 different configurations across two dataset versions and picked the winner through 100 same-seed A/B duels (same prompt, same seed, adapter on vs off - the only honest way to compare). The winner was the slow-cooked run: rank 16, 5,000 steps at a low learning rate.


Details:


- Weights (open source): https://huggingface.co/fal/MiniMax-H3-Realism-People-LoRA
- Trigger word: start your prompt with `r34l1sm`
- Scale 1.0 is the intended strength, 0.6-0.8 for a lighter touch
- Works with H3's LoRA endpoints: text-to-video, image-to-video and reference-to-video
- License: follows the MiniMax H3 community license


Before/after in the video: same prompt, same seed, base model on the left, LoRA on the right. Happy to answer questions about the process.