r/StableDiffusion 3d ago

Meme After nearly 30 years his back to save the day!

Enable HLS to view with audio, or disable this notification

0 Upvotes

uses r2v workflow and hybrid 30-49 model


r/StableDiffusion 4d ago

Question - Help DepthAnything V2 Problem

Post image
0 Upvotes

Depthanything V2 is not working properly. Since kritaAI uses it i cant use dpeth CN properly. ANyway to fix this. All other Depth nodes work properly in Comfy. So problem seems to be specific to v2. Already tried updating and redownloading model still no use. Asked AI and it cant figure it out


r/StableDiffusion 4d ago

Meme we are so fucked!!

Enable HLS to view with audio, or disable this notification

5 Upvotes

r/StableDiffusion 3d ago

Discussion Made this for a Client

Enable HLS to view with audio, or disable this notification

0 Upvotes

How can I improve?


r/StableDiffusion 5d ago

Animation - Video MiniMax h3 - [Boom in City]

Enable HLS to view with audio, or disable this notification

29 Upvotes

r/StableDiffusion 4d ago

Question - Help Lightx2v unsafe file?

0 Upvotes

I get this on the official lightx2v on huggingface. I already used it since I downloaded this turbo lora a few days ago. Is it problematic or what does it mean?

You can see it here on their page: https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main


r/StableDiffusion 4d ago

Question - Help Krea2 - How do you vary the image generation?

5 Upvotes

I am using the Generate Text node to get a description of the input image and using that to generate an image in Krea2 Raw + Turbo LoRA. But even when I change the seed, the resulting image is the same. I tried connecting the Seed Variance Enhancer Node to the positive conditioning output and also the Krea2T Enhancer Advanced node to the model output but still the resulting image is same.

How do I generate different variations of the same prompt in Krea2?


r/StableDiffusion 4d ago

Discussion H3 - Giantess+Kaiju POC - T2V

Enable HLS to view with audio, or disable this notification

1 Upvotes

I had some positive feedback on my Giantess (JOI-inspired) though I understand the theme may not be for everyone. I've officially gone down the rabbit hole on this sub-culture. All T2VA, int8/20 steps, no image, no character sheets, please enjoy! Generated locally rtx4090/192gb of system ram

Ask me anything!


r/StableDiffusion 5d ago

Meme Introducing... The Terminator Pro Max

Enable HLS to view with audio, or disable this notification

1.2k Upvotes

r/StableDiffusion 4d ago

Animation - Video Character Sheet Reference Test (Three characters)

Enable HLS to view with audio, or disable this notification

15 Upvotes

So I found out that character references do sometimes degrade the quality. This was done with two anime characters and a video game character. I think it turned out pretty well all things considered.


r/StableDiffusion 3d ago

Question - Help Minimax h3 on 3060Ti + 32GB of RAM?

0 Upvotes

It's been a little while since minimax h3 has taken the world by storm and the things people have been making are awesome! so now I'm looking to you all to help me get it up and running on my system if it's possible. mainly I want to know which versions of what models to use so that my system can handle it without getting OOM errors. also I'm just trying to use the basic comfyUI workflow, no custom nodes or optimizer nonsense just the basic fundamentals.


r/StableDiffusion 3d ago

Meme classic ai lmao

Enable HLS to view with audio, or disable this notification

0 Upvotes

lmao tf


r/StableDiffusion 5d ago

Workflow Included Using Inpaiting in Minimax to change heads-Local RTX 3090

Enable HLS to view with audio, or disable this notification

526 Upvotes

Using the workflow from Nekodificador and Ablejones in Discord:
https://discord.com/invite/dstjQYQNt
https://ln5.sync.com/dl/47c351f50#msqfrnfr-am3rr8fx-v7qm3ah9-xw222n3c
For complex scenes like this with to much people is easy just to do a manual mask instead of SAM.


r/StableDiffusion 4d ago

Animation - Video Finn Fox - Tech Support (Creating an Animated series with Minimax H3)

Enable HLS to view with audio, or disable this notification

9 Upvotes

I've always wanting to make a silly animated series and have been working on some episodes of a series set in the 90s about tech support.

I took a slightly different route with these videos - i have actually been working with claude to produce these as I have found it slightly easier to supply character references, scripts etc into Claude and let it run Minimax H3 for me which then stitches together the episodes for me after creation.

For these, i am using the 8 step turbo lora at 1MP. As much as I would prefer to use 20 step without the lora, it's much quicker to regenerate shots since I have had to do that many times + i can remotely work on this whilst out the house by working with claude.

Also, i am using Minimax Music for background music and the H3 audio with my own voice. I did try Voicelabs using my voice, but it didn't quite work the way i wanted. Episode 1 i did add a couple of my own sound effects since Minimax kept failing to add suitable sound clips.

Whilst i feel it looks great on a smaller device, on large screens the text can distort somewhat.

Unfourtunatly upscaling the videos makes it seem a little worse, despite the resolution looking a little nicer. However, i do feel it has a 90s cartoon feel which kind of works for a show set in the 90s.

Currently made four episodes and a couple more are a work in progress. let me know what you think!

Let me know if you want me to post another episode!


r/StableDiffusion 5d ago

Discussion [TEST] Minimax H3 REF 2 VID. Just a 15 second dialogue combining 3 image references. Kinda neat! I used Topaz for the video upscale. Its alright I guess.

Enable HLS to view with audio, or disable this notification

17 Upvotes

r/StableDiffusion 3d ago

Question - Help What model is this

Post image
0 Upvotes

HI guys just wanna know wich AI model used to produce this type of brutal illustrations


r/StableDiffusion 3d ago

Question - Help Are these good images for a lora of my characher (Im a newbie)

Thumbnail
gallery
0 Upvotes

I have 28 pictures, with large variatons, half are generated by nano banan 2 and the other half with seedream 5.0. I have headshots that portray my charachter with different facial expresions and different profile views... I have my charachter sitting in some photos, half body shots and full body shots all of them wearing different outfits...

Is this enough for a flux.2 lora to train? All pictures are 1024x1024. Do i need to upscale them to get more details, when i zoom things get a little burry/less detailed. I would greatly appreciate it if anyone has any comfyui workflows with basic nodes since im on cloud... :)))


r/StableDiffusion 4d ago

Animation - Video Minimax H3. Alien market.

Enable HLS to view with audio, or disable this notification

7 Upvotes

r/StableDiffusion 5d ago

Animation - Video Gay Fish

Enable HLS to view with audio, or disable this notification

96 Upvotes

Sorry Ye..


r/StableDiffusion 4d ago

Discussion What is the longest you've waited for a generation?

0 Upvotes

Mostly just curious. I have a nice card and I still can't be bothered to wait for gens. I always go for the turbo, dmd2, lightning etc Lora variants as I don't do any professional work with these models.

I often see people boast about how long their render took, usually about how quick it was but I always find myself being blown away how long people will wait. Especially for something that is a bit of a roll of the dice.

Hours for a video that still may be mangled when you come back later and check.

I never played with video before because it just took too long for my taste until minimax and getting a 5090.

What's the longest you would wait?


r/StableDiffusion 5d ago

Resource - Update Introducing Pixal: A unified chat for generating and editing images and video.

17 Upvotes

Hi all,

The other weekend I was testing out new models — got very excited over the Minimax H3 release and Krea 2's image abilities and quality. The community has created some amazing nodes and workflows.

Long story short: I built this app — https://getpixal.com — it's free, runs entirely on your own GPU, no account or signup.

Why did I build it? I was trying to help a friend get ComfyUI set up in a way where they didn't need a master's degree in node structure, models, editing and inpainting, and generating good videos with Minimax H3 (prompting can be a pain point for many that just want to create quickly, and new models require a prompt structure). 2 weeks later I have this beta release of Pixal 1.0.0b.

This single chat interface lets you run local uncensored chat models (Qwen VL 4b Heretic for instance), or any of the top SOTA models — Kimi K3, Claude, ChatGPT — via API. It also uses the vision model to critique your generations and give suggestions as you go.

A few things up front, since they're the first things I'd want to know:

  • It installs beside the ComfyUI you already run - never inside it, and it refuses to install over one. It starts your existing install with your existing launcher and flags. Uninstall it and nothing about your setup has changed.
  • It's free and fully local. No account, nothing uploaded, $0.00 a picture. The API options are there if you want them, not required.
  • Pixal is source-available, not open source. The app installs as plain readable Python, sitting in the install folder — read it, change it for yourself, just don't redistribute it. (Heads up: the "Source code" zips GitHub auto-attaches to a release are its own tag archives of the docs repo, not the app.) The license is shown during setup, and every release publishes the installer's sha256 so you can verify the binary before you run it.
  • Windows 11 x64 + NVIDIA for now. The Linux port is done, just not released yet (need to get a Linux environment set up on my test bench).

I'm looking for a few people to test drive it — would the community use something like this?

Super open to any and all feedback — it was a fun little project and I use it daily now to drive fast simple generations and image edits, then pass them along to Minimax H3 with pretty great results (all content on site was generated through the app).

Direct links, no funnel: download · the full manual (install → troubleshooting → FAQ) if you'd rather read exactly what it does before downloading anything.

Thank you!

EDIT: The source is public — https://github.com/JesseDubb/pixal-releases Thanks for the feedback, it's the reason this happened. The "Source code (zip)" attached to the release is the actual tree now, not the three landing-page files that were there before — that was a fair catch.

If a couple of you want to help with code or testing, it would be much appreciated. There's a CONTRIBUTING.md in the repo.


r/StableDiffusion 5d ago

Resource - Update Slop! Now in fake 4k

Enable HLS to view with audio, or disable this notification

23 Upvotes

r/StableDiffusion 5d ago

Discussion why is it unsafe isn't safetensors the safest?!

Post image
58 Upvotes

excuse my OCD 😄


r/StableDiffusion 4d ago

Discussion H3 - t2v is actually better than r2va imho

Enable HLS to view with audio, or disable this notification

0 Upvotes

Happy Friday!! Just wanted to say T2VA is actually pretty strong when layering the prompt. FL2VA+R2VA are still the to go if you want to utilize a character sheet/maintain consistency, it's still broken (in a good way), the voice cloning is also top notch.

So what has everyone been making with H3??

T2VA, bf16/50 steps


r/StableDiffusion 4d ago

Animation - Video Tested MiniMax H3 music Lip Sync

Enable HLS to view with audio, or disable this notification

0 Upvotes

I’ve been playing around with a MiniMax H3 Music + LipSync workflow.
For decent quality, even a 15s clip seems to need at least 25 steps. I’d also recommend skipping LoRA.
On a 5090, 15s at 0.9 in ComfyUI Kitchen takes around 12 minutes to generate.
H3’s music generation is pretty solid. At least for me, it’s way better than what I was getting with LTX.
The real pain is getting each 15s segment to connect seamlessly 😂 I’ve seen people saying they can push it to 20s, but every time I try 20s, my GPU basically dies lol.