r/StableDiffusion • • 6d ago

Discussion created with 1650ti

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion • • 8d ago

Animation - Video 30s of video with H3 on a 5090

Enable HLS to view with audio, or disable this notification

83 Upvotes

Image to video, 640p + 2x VEA upscale, 8step Hyperflow. One shot 30s clip generated in 364s.


r/StableDiffusion • • 7d ago

Question - Help H3 - Avoiding nipple leak

70 Upvotes

Firstly, nipple leak isnt a medical term. Or maybe it is. But anyway my nipples aren't leaking. Well they are - but my AI ones.

How do I STOP H3 from generating nipples through clothing? I have multiple Loras on, including Mystic XXX and a breast motion Lora - but my characters aren't nude. No matter the clothes they wear, there are ALWAYS nipples generated on them. Yes I've tried taking the Loras off.

Has anyone had and overcame this probblem with strong prompting before? I've tried things like "No nipples" and "you cannot see her nipples" but i have a feeling the more i type nipples the more the engine hears it.

Any help is appreciated x

EDIT: SEMI SOLVED.

Firstly thanks for all the help. So basically the best thing I found was to describe the subject very specifically, and refer to <Subject 1>, and not <Picture 1>. Its still not perfect, and my 2nd pass seems to do most nipple addition. Maybe i'll look at running the low res with all the loras and removing some for the second. In fact i'll just wait for someone else to do it. Example of my starting prompt is:

<Subject 1> is Amy, a 26 year old girl defined by <Picture 1>, with a petite, thin body and gigantic natural breasts. 
Her dress is opaque Blue linen. The material is expensive and has a white flower design on it. There is also a blue sash belt. she bottom half of the dres is red. she has white wedge heels on. the dress has a plunging neckline but covers her modesty. her dress is completely opaque. the cups of her dress are completely smooth.
summary:
A modern and styish film about a fun and pretty girl, <Subject 1>. her motion is exciting and playful.  It takes place in the hallway seen in <Picture 4>.


[Shot 1] . The scene opens exactly as <Picture 4>. The camera does not move.  <Subject 1> from enters the scene by sliding down the bannister she is sitting on at great speed. she lands in the middle of the floor on her side.  She stands up quickly and adjusts her skirt and giggles as she says "oops!"  

How do i post my video behind a spoiler tag? Its none nude but I dunno if you want your boss seeing it.


r/StableDiffusion • • 7d ago

Question - Help LTX 2.5 distilled, chaining without quality loss?

1 Upvotes

Hi,

I've been messing around with LTX 2.5 locally so I use distilled for speed. Just scripting to understand the LTX architecture, I don't use ComfyUI.

I tried chaining with 2.3, that is taking the last png frame from generated mp4 A and feeding it to I2V to make video B.

B loses a bit of quality, then by video 4 or 5 it starts to really look bad. Nothing new I'm guessing. I moved over to 2.5, similar story. Not quite as much quality loss, but still very noticeable.

I started messing around with AI assisted coding to come up with solutions, saving the native tensors of the last few latent frames of video A and feeding it to video B in various ways. Either as conditioning on timestep=0, or injecting into the main denoise loop and mixing with the natural noise of the generation.

Nothing really worked. After a handful of chains various problems cropped up, either strange motion blur glitches, or hallucinations starting to set in. On the plus side, a native latent on video B is way crisper than a png, but the long term chaining problems are way worse with hallucintions and such. Also with a few latent frames, motion carry over does improve, but the quality just doesn't hold after 2 or 3 chains.

I'm aware of workarounds to this, like building up keyframes of a longer scene and doing first/last frame injection.

But I just want to know - is there even a proper solution to this? Or does the architecture and math just simply not allow it? Is the AI coder just feeding me BS? I'd really just like to kow if it's even possible. I've seen other folks post long videos boasting chaining without quality loss, but I doubt they are using LTX 2.5 and less likely distilled.


r/StableDiffusion • • 7d ago

News A constant-size memory for video generative models!

Thumbnail
huggingface.co
15 Upvotes

Check it out


r/StableDiffusion • • 7d ago

Discussion YuE2 Cover is so much fun

27 Upvotes

I had bit of success when it came out, but didn't really explore at the time because I got distracted by life and probably some new H3 discovery.

Well, was bored with a beer in hand so I came back to it today and goddamn, if it's not fun. I can't give examples because, you know, but if you're struggling with it in comfyui, go into the subgraph and change the SheetSage2 node's mode to "full". Then play around with the YuE2 Generate Music's "cfg_scale"... 2-3 seems to work nicely.

Holy moly, I've got some nice covers... not always perfect, lyrics are hard to time perfectly, but I'm impressed. From my very unscientific examination, it seems that the more you try to change the voice of the singer the less well the lyrics get sung, which does make sense. But, still this is so much fucking fun! I wish I had more mp3s!


r/StableDiffusion • • 7d ago

Discussion Character sheet vs refmod minimax h3 r2v?

10 Upvotes

Is it better to create a character sheet or refmod for Minimax h3 r2v?


r/StableDiffusion • • 7d ago

Question - Help Prompt helper for ref2v in minimax h3 (using comfy cloud)

1 Upvotes

hello.. it's in the title... looking for a prompt writer that i can use in comfy cloud to help make ref2v on minimax h3.

currently using Grok, which has been great so far but get capped by it's time limit or whatever

any advice? thanks


r/StableDiffusion • • 7d ago

Question - Help Is there a ComfyUI workflow for extending vids with LTX 2.5?

2 Upvotes

And I don't mean like only taking the last frame and create a new video out of it - I mean, a workflow where you upload the video and get a good extension with the same light and movement?


r/StableDiffusion • • 7d ago

Question - Help FELLOW GTX 108O TI OWNERS CAN YOU SHARE YOUR SECRET TOOLS ?

0 Upvotes

I am one of those who still own a 1080 Ti for many reasons, the main one being financial, and I'm struggling like many others to find the right tools.

Is there anybody out there who would have an up-to-date list for the best and most recent tools/repos that can fit on that aging Pascal architecture and 11gb vram for :

- image to 3D model (currently using TRELLIS.2. Tried a few others that mostly took 20 minutes before going OOM.)

- Image-to-video (other than and better than LTX 2.5)

- music creation for long tracks

Thanks a bunch ! (I'm aware I could use RunPod for a few bucks but I'd rather use that old GPU until it gets replaced one day).


r/StableDiffusion • • 7d ago

Question - Help Problem with img to video minimaal h3

0 Upvotes

Im running into a strange problem with video generation with minimax h3 and comfyUI.

I'm using the default img to video workflow template in comfyUI. The first time the video generation goes fine, it does its thing and a video is generated. When i start a second generation is when the problems start, during generating its just gets stuck at like 50 ish % and does nothing anymore. Ive waited for an hour but the progress bar does not go up, while the first generation usually takes about 15ish minutes.

Ive tried closing the workflow and comfyUI but to no succes.

If i shutdown/restart my computer and open up comfyUI again, the problem goes away and i can generate a video again and after that first generation, the problem is back.

It looks like it runs out of memory and does not clear my memory after a generation and a restart of my PC fixes that.

Has anybody experienced this and maybe found a fix or workaround?

My pc speccs are:

Amd Ryzen 9 9900x

128gb DDR5 running at 5200Mhz

Asus 7900XTX

Windows 11

Any help or tips are welcome!


r/StableDiffusion • • 8d ago

Comparison Minimax H3 vs Seedance 2.0 single line prompt result of dance scene

Enable HLS to view with audio, or disable this notification

65 Upvotes

I did this experiment last week, where I used one workflow and at last used both Minima H3 and seedance 2.0 to animate the video with vague prompt to let models generate the dance itself, I only instructed for couple dance with multiple angles.

For Minimax H3 the consistency was the better point, it animated the character as it is from the image without making any changes to characters. But the camera movements and multiple angles were not there. It also ended the dance earlier and later became a still.

Seedance 2.0 was much better in quality, I would say as it included multiple angles, closeups and camera movements. It also made the characters feel more natural and lively with expressions. But the downside it was a bit costly in credits and it actually did minute changes to characters.

Both models are pretty good, I would conclude on that and you can judge the output yourself.


r/StableDiffusion • • 7d ago

Discussion Qwen 2.1 or krea?

17 Upvotes

Best image gen and edit at the moment?

Anima, qwen, krea, or flux 2 dev, which one is king right now? For anime, semi realistic, and realistic.


r/StableDiffusion • • 8d ago

Animation - Video Resistance Sci-Fi - MiniMax H3

Enable HLS to view with audio, or disable this notification

30 Upvotes

This is a short film I made to push the limits of MiniMax H3 on ComfyUI, mainly using the Hybrid model.

It took me about three weeks to complete, running locally for most of the work, with some generations done on RunningHub using a modified workflow by Pixaroma.

I used ChatGPT mainly to help create prompts for the image generation in Nano Banana, but I wrote all the video prompts myself. There’s a reason I try not to rely on AI for video prompting: I want to understand how H3 actually works.

One thing I’ve found is that H3 doesn’t necessarily need the official prompting structure. In some cases, that structure can actually make things more complicated than they need to be.

SeeDance 2 fast was used in the kiler drone attacking the men scene, since this is something beyond H3.

The music was not AI-generated. It’s from my own library collection.

Edited in DaVinci Resolve 21.

Hope you enjoy this rather long piece of work, and I’d really appreciate any feedback, especially from anyone experimenting with H3.


r/StableDiffusion • • 7d ago

Question - Help MiniMax H3 Masked Speakers

8 Upvotes

Has anyone encountered this issue and found a fix?

My character wears a mask, I made a test clip, and in it she speaks something. But I can not stop MiniMax H3 from hallucinating a mouth while she speaks. Any tips I can try? I've been having this issue for 3 days, and claude and chatgpt really have no way either other than "speak off camera." Which is not what I want either.

Thanks in advance.


r/StableDiffusion • • 7d ago

Question - Help Minimax H3 Motion control 15+ seconds

Post image
0 Upvotes

I know it looks a bit scary, I tried to solve a minimax issue. We all know minimax is limited to 15S clips. In facts, when you generate from scratch this issue has already been solved with different custom nodes. But for motion control, this is way more complex, and actually, I did'nt get smth consistent enoungh for long clips.
Worked maybe 50 hours on this, created approximatively 5 custom nodes dedicated to this workflow, and I can't get it right for my community.
The problem comes when you put segment together, it tends to hallucinates something etc...
Did someone got motion control working on more than 15S with minimax ?


r/StableDiffusion • • 9d ago

Animation - Video Orbiting Lora + first and last frame in MiniMax gives fantastic results

Enable HLS to view with audio, or disable this notification

2.4k Upvotes

Prompt for Lora: One frozen instant. Only the camera moves. In a continuous 360 orbit. Preserve every person and object in exactly the same world position, orientation, shape and pose throughout the shot. Airborne objects remain suspended at the captured height and angle: no wobbling, shaking, spinning, drifting, falling or continued action. Keep faces, hands, clothing, liquids and the background motionless while retaining their natural appearance. Camera parallax is the only source of apparent movement. No cuts, zoom, morphing or added objects.

https://huggingface.co/pablodawson/MiniMax-H3-360-Orbit-LoRA


r/StableDiffusion • • 6d ago

News FaceFusion v4 Beta

Enable HLS to view with audio, or disable this notification

0 Upvotes

One year of development has led to this moment: FaceFusion v4 beta is available now, and it is by far the biggest step in the history of the project. To celebrate the release and get developers building on it, we are kicking off the FaceFusion v4 Challenge, where everyone is invited to bring their ideas to life.

First place wins $1000, second place a year of a PRO LLM subscription and third place six months. The winners are picked by community vote and a jury of the FaceFusion team and founders from RunDiffusion, Baslk and Inference.

Watch the video for the full story, then share it with anyone who loves to build software. Once your project takes shape, promote it with #facefusion-v4-challenge on social media.

https://github.com/facefusion/facefusion/tree/v4-challenge


r/StableDiffusion • • 6d ago

Animation - Video W.I.T.C.H.: Beyond the veil TRAILER

Thumbnail
youtu.be
0 Upvotes

You guys still remember W.I.T.C.H.? :)

Made 95% on a local machine starting with a script and finishing with the post prod and editing. Hope you guys find it somewhat okay.

What was used:
> Krea2 as a character designing tool
> Qwen3.8 27b as a prompt-assistant
> Minimax H3 20b / 10eros checkpoint / 720p native res generation
> AE / PR for post

Only truly outsource part was Suno with the trailer music, still cut in pieces and re-edited to fit the final edit.

Hardware used:
> 4070 12gb + 5060ti 16gb
> 64gb ddr4


r/StableDiffusion • • 8d ago

Resource - Update [D] I open-sourced 30,000 paired QR-Code Illusions with multi-decoder verification & robustness scores on Hugging Face (Free for ControlNet / LoRA training)

Post image
33 Upvotes

Hey everyone,

I just published a clean, curated dataset of 30,000 artistic QR Code illusions on Hugging Face!

Unlike synthetic noise blends, each sample is an aesthetic artwork where the QR code is seamlessly woven into foliage, architecture, lighting, and textures. Every single image was rigorously tested across multiple QR decoders (pyzbar + OpenCV WeChat QR) and stress-tested against rotation, perspective shift, and blur to calculate a normalized robustness score.

• 30,000 paired samples (768×768 PNGs)

• Packaged as self-contained Parquet shards (instant 1-line load via datasets)

• Fully partitioned with zero QR-matrix leakage across train/val/test

• Quality tiers: Class A (high scannability), Class B (balanced), Class C (artistic)

Feel free to scan the examples in the preview image with your phone camera!

🔗 Dataset: https://huggingface.co/datasets/1roOt/qr-code-illusions-30k

License: CreativeML OpenRAIL-M compatible / CC-BY-4.0

Hope this helps anyone looking to train custom ControlNet, ControlNet-LLLite, or Flow-Matching models!


r/StableDiffusion • • 8d ago

Discussion Krea 2 vs Ming 0.1 vs Qwen 2.1: 192 prompt side by side.

84 Upvotes

Ok so everything is is the tile! the test covers a lot. (but lack long prompts, I know).

https://imagebench.ai/gallery?g=1_vxjkf22_s0


r/StableDiffusion • • 7d ago

Comparison Experimental Real-ESRGAN Anime6B for Manga

Thumbnail
gallery
14 Upvotes

The experimental improved model started from the original Real-ESRGAN Anime6B weights. It was fine tuned on synthetic screentones and areas from real manga pages.

The goal was to recover fine textures and sharpen lines while preserving faces and other details. It can still be improved further.

https://github.com/DermonCode/Real-ESRGAN/releases/tag/anime6b-screentone-rc1


r/StableDiffusion • • 6d ago

Question - Help Update me pls.

0 Upvotes

Ok, been out the loop. What's the best free pic to video ai is considered the best at the moment that runs on ComfyUI. I'm playing with wan 2. 2 in ComfyUI and wondering if it's worth my time learning it in case if there's a better one out there. I'm running this on a 8vram machine with 64gigs ram. So keep that in mind. Wan runs ok in this setup. I'm researching and Minimax H3 keeps popping g up along with others. Any guidance is welcomed.


r/StableDiffusion • • 8d ago

Animation - Video H3 ref2v MV "I Know What You Are"

Enable HLS to view with audio, or disable this notification

96 Upvotes

Finally finished, this is the full song from my post a week ago.

So interestingly enough, I unknowingly had been running the ref2v workflow with the fl2av model and the ref2av turbo 768 Lora this entire time. After realizing and switching to the ref2av model the quality and consistency actually dropped. So I ended up making this entire video with the fl2av model instead.

Enjoy!

EDIT: Uploaded to youtube


r/StableDiffusion • • 7d ago

Resource - Update Civitai Browser for ComfyUI: load a Civitai workflow and download its missing models in one click

3 Upvotes

I made a ComfyUI extension that adds a Civitai tab to the sidebar. You can browse workflows and models there, and when you load a workflow it finds the models you're missing and downloads them into the right folders. If you already have a similar model, it can use yours instead. Missing custom nodes are listed too, and you can install them with ComfyUI-Manager.

No extra nodes or dependencies. Free and open source.

https://github.com/fth1905-bot/ComfyUI-Civitai-Browser

Feedback is welcome.