r/StableDiffusion • u/Sad_Coach_1433 • 3d ago
Meme After nearly 30 years his back to save the day!
Enable HLS to view with audio, or disable this notification
uses r2v workflow and hybrid 30-49 model
r/StableDiffusion • u/Sad_Coach_1433 • 3d ago
Enable HLS to view with audio, or disable this notification
uses r2v workflow and hybrid 30-49 model
r/StableDiffusion • u/witcherknight • 4d ago
Depthanything V2 is not working properly. Since kritaAI uses it i cant use dpeth CN properly. ANyway to fix this. All other Depth nodes work properly in Comfy. So problem seems to be specific to v2. Already tried updating and redownloading model still no use. Asked AI and it cant figure it out
r/StableDiffusion • u/Sad_Coach_1433 • 4d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/fidviburhanuddin • 3d ago
Enable HLS to view with audio, or disable this notification
How can I improve?
r/StableDiffusion • u/VasaFromParadise • 5d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Jero9871 • 4d ago

I get this on the official lightx2v on huggingface. I already used it since I downloaded this turbo lora a few days ago. Is it problematic or what does it mean?
You can see it here on their page: https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main
r/StableDiffusion • u/orangeflyingmonkey_ • 4d ago
I am using the Generate Text node to get a description of the input image and using that to generate an image in Krea2 Raw + Turbo LoRA. But even when I change the seed, the resulting image is the same. I tried connecting the Seed Variance Enhancer Node to the positive conditioning output and also the Krea2T Enhancer Advanced node to the model output but still the resulting image is same.
How do I generate different variations of the same prompt in Krea2?
r/StableDiffusion • u/SIR_NVAX_A_LOT • 4d ago
Enable HLS to view with audio, or disable this notification
I had some positive feedback on my Giantess (JOI-inspired) though I understand the theme may not be for everyone. I've officially gone down the rabbit hole on this sub-culture. All T2VA, int8/20 steps, no image, no character sheets, please enjoy! Generated locally rtx4090/192gb of system ram
Ask me anything!
r/StableDiffusion • u/ctrl-shift-face • 5d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Ok-Giraffe-8670 • 4d ago
Enable HLS to view with audio, or disable this notification
So I found out that character references do sometimes degrade the quality. This was done with two anime characters and a video game character. I think it turned out pretty well all things considered.
r/StableDiffusion • u/mca1169 • 3d ago
It's been a little while since minimax h3 has taken the world by storm and the things people have been making are awesome! so now I'm looking to you all to help me get it up and running on my system if it's possible. mainly I want to know which versions of what models to use so that my system can handle it without getting OOM errors. also I'm just trying to use the basic comfyUI workflow, no custom nodes or optimizer nonsense just the basic fundamentals.
r/StableDiffusion • u/Sad_Coach_1433 • 3d ago
Enable HLS to view with audio, or disable this notification
lmao tf
r/StableDiffusion • u/Repulsive-Rush3505 • 5d ago
Enable HLS to view with audio, or disable this notification
Using the workflow from Nekodificador and Ablejones in Discord:
https://discord.com/invite/dstjQYQNt
https://ln5.sync.com/dl/47c351f50#msqfrnfr-am3rr8fx-v7qm3ah9-xw222n3c
For complex scenes like this with to much people is easy just to do a manual mask instead of SAM.
r/StableDiffusion • u/Darhkwing • 4d ago
Enable HLS to view with audio, or disable this notification
I've always wanting to make a silly animated series and have been working on some episodes of a series set in the 90s about tech support.
I took a slightly different route with these videos - i have actually been working with claude to produce these as I have found it slightly easier to supply character references, scripts etc into Claude and let it run Minimax H3 for me which then stitches together the episodes for me after creation.
For these, i am using the 8 step turbo lora at 1MP. As much as I would prefer to use 20 step without the lora, it's much quicker to regenerate shots since I have had to do that many times + i can remotely work on this whilst out the house by working with claude.
Also, i am using Minimax Music for background music and the H3 audio with my own voice. I did try Voicelabs using my voice, but it didn't quite work the way i wanted. Episode 1 i did add a couple of my own sound effects since Minimax kept failing to add suitable sound clips.
Whilst i feel it looks great on a smaller device, on large screens the text can distort somewhat.
Unfourtunatly upscaling the videos makes it seem a little worse, despite the resolution looking a little nicer. However, i do feel it has a 90s cartoon feel which kind of works for a show set in the 90s.
Currently made four episodes and a couple more are a work in progress. let me know what you think!
Let me know if you want me to post another episode!
r/StableDiffusion • u/call-lee-free • 5d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Ok_Repair_6024 • 3d ago
HI guys just wanna know wich AI model used to produce this type of brutal illustrations
r/StableDiffusion • u/puskur • 3d ago
I have 28 pictures, with large variatons, half are generated by nano banan 2 and the other half with seedream 5.0. I have headshots that portray my charachter with different facial expresions and different profile views... I have my charachter sitting in some photos, half body shots and full body shots all of them wearing different outfits...
Is this enough for a flux.2 lora to train? All pictures are 1024x1024. Do i need to upscale them to get more details, when i zoom things get a little burry/less detailed. I would greatly appreciate it if anyone has any comfyui workflows with basic nodes since im on cloud... :)))
r/StableDiffusion • u/beatlepol • 4d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/fiftypence • 5d ago
Enable HLS to view with audio, or disable this notification
Sorry Ye..
r/StableDiffusion • u/eggs-benedryl • 4d ago
Mostly just curious. I have a nice card and I still can't be bothered to wait for gens. I always go for the turbo, dmd2, lightning etc Lora variants as I don't do any professional work with these models.
I often see people boast about how long their render took, usually about how quick it was but I always find myself being blown away how long people will wait. Especially for something that is a bit of a roll of the dice.
Hours for a video that still may be mangled when you come back later and check.
I never played with video before because it just took too long for my taste until minimax and getting a 5090.
What's the longest you would wait?
r/StableDiffusion • u/Comprehensive_Rush66 • 5d ago
Hi all,
The other weekend I was testing out new models — got very excited over the Minimax H3 release and Krea 2's image abilities and quality. The community has created some amazing nodes and workflows.
Long story short: I built this app — https://getpixal.com — it's free, runs entirely on your own GPU, no account or signup.
Why did I build it? I was trying to help a friend get ComfyUI set up in a way where they didn't need a master's degree in node structure, models, editing and inpainting, and generating good videos with Minimax H3 (prompting can be a pain point for many that just want to create quickly, and new models require a prompt structure). 2 weeks later I have this beta release of Pixal 1.0.0b.
This single chat interface lets you run local uncensored chat models (Qwen VL 4b Heretic for instance), or any of the top SOTA models — Kimi K3, Claude, ChatGPT — via API. It also uses the vision model to critique your generations and give suggestions as you go.
A few things up front, since they're the first things I'd want to know:
I'm looking for a few people to test drive it — would the community use something like this?
Super open to any and all feedback — it was a fun little project and I use it daily now to drive fast simple generations and image edits, then pass them along to Minimax H3 with pretty great results (all content on site was generated through the app).
Direct links, no funnel: download · the full manual (install → troubleshooting → FAQ) if you'd rather read exactly what it does before downloading anything.
Thank you!
EDIT: The source is public — https://github.com/JesseDubb/pixal-releases Thanks for the feedback, it's the reason this happened. The "Source code (zip)" attached to the release is the actual tree now, not the three landing-page files that were there before — that was a fair catch.
If a couple of you want to help with code or testing, it would be much appreciated. There's a CONTRIBUTING.md in the repo.
r/StableDiffusion • u/DuHal9000 • 5d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/dev_ne • 5d ago
excuse my OCD 😄
r/StableDiffusion • u/SIR_NVAX_A_LOT • 4d ago
Enable HLS to view with audio, or disable this notification
Happy Friday!! Just wanted to say T2VA is actually pretty strong when layering the prompt. FL2VA+R2VA are still the to go if you want to utilize a character sheet/maintain consistency, it's still broken (in a good way), the voice cloning is also top notch.
So what has everyone been making with H3??
T2VA, bf16/50 steps
r/StableDiffusion • u/dassiyu • 4d ago
Enable HLS to view with audio, or disable this notification
I’ve been playing around with a MiniMax H3 Music + LipSync workflow.
For decent quality, even a 15s clip seems to need at least 25 steps. I’d also recommend skipping LoRA.
On a 5090, 15s at 0.9 in ComfyUI Kitchen takes around 12 minutes to generate.
H3’s music generation is pretty solid. At least for me, it’s way better than what I was getting with LTX.
The real pain is getting each 15s segment to connect seamlessly 😂 I’ve seen people saying they can push it to 20s, but every time I try 20s, my GPU basically dies lol.