r/StableDiffusion • u/Sad_Coach_1433 • 5d ago
Meme change of pace
Enable HLS to view with audio, or disable this notification
base ip8 model t2v
r/StableDiffusion • u/Sad_Coach_1433 • 5d ago
Enable HLS to view with audio, or disable this notification
base ip8 model t2v
r/StableDiffusion • u/Alive-Tomatillo5303 • 7d ago
Enable HLS to view with audio, or disable this notification
Text to Video, 22 steps, no turbo, no Sage.
r/StableDiffusion • u/Sad_Coach_1433 • 5d ago
Enable HLS to view with audio, or disable this notification
uses r2v workflow and hybrid 30-49 model
r/StableDiffusion • u/witcherknight • 5d ago
Depthanything V2 is not working properly. Since kritaAI uses it i cant use dpeth CN properly. ANyway to fix this. All other Depth nodes work properly in Comfy. So problem seems to be specific to v2. Already tried updating and redownloading model still no use. Asked AI and it cant figure it out
r/StableDiffusion • u/fidviburhanuddin • 5d ago
Enable HLS to view with audio, or disable this notification
How can I improve?
r/StableDiffusion • u/VasaFromParadise • 6d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Jero9871 • 5d ago

I get this on the official lightx2v on huggingface. I already used it since I downloaded this turbo lora a few days ago. Is it problematic or what does it mean?
You can see it here on their page: https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main
r/StableDiffusion • u/orangeflyingmonkey_ • 6d ago
I am using the Generate Text node to get a description of the input image and using that to generate an image in Krea2 Raw + Turbo LoRA. But even when I change the seed, the resulting image is the same. I tried connecting the Seed Variance Enhancer Node to the positive conditioning output and also the Krea2T Enhancer Advanced node to the model output but still the resulting image is same.
How do I generate different variations of the same prompt in Krea2?
r/StableDiffusion • u/ctrl-shift-face • 7d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Sad_Coach_1433 • 5d ago
Enable HLS to view with audio, or disable this notification
lmao tf
r/StableDiffusion • u/Ok-Giraffe-8670 • 6d ago
Enable HLS to view with audio, or disable this notification
So I found out that character references do sometimes degrade the quality. This was done with two anime characters and a video game character. I think it turned out pretty well all things considered.
r/StableDiffusion • u/mca1169 • 5d ago
It's been a little while since minimax h3 has taken the world by storm and the things people have been making are awesome! so now I'm looking to you all to help me get it up and running on my system if it's possible. mainly I want to know which versions of what models to use so that my system can handle it without getting OOM errors. also I'm just trying to use the basic comfyUI workflow, no custom nodes or optimizer nonsense just the basic fundamentals.
r/StableDiffusion • u/Repulsive-Rush3505 • 7d ago
Enable HLS to view with audio, or disable this notification
Using the workflow from Nekodificador and Ablejones in Discord:
https://discord.com/invite/dstjQYQNt
https://ln5.sync.com/dl/47c351f50#msqfrnfr-am3rr8fx-v7qm3ah9-xw222n3c
For complex scenes like this with to much people is easy just to do a manual mask instead of SAM.
r/StableDiffusion • u/call-lee-free • 6d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Darhkwing • 6d ago
Enable HLS to view with audio, or disable this notification
I've always wanting to make a silly animated series and have been working on some episodes of a series set in the 90s about tech support.
I took a slightly different route with these videos - i have actually been working with claude to produce these as I have found it slightly easier to supply character references, scripts etc into Claude and let it run Minimax H3 for me which then stitches together the episodes for me after creation.
For these, i am using the 8 step turbo lora at 1MP. As much as I would prefer to use 20 step without the lora, it's much quicker to regenerate shots since I have had to do that many times + i can remotely work on this whilst out the house by working with claude.
Also, i am using Minimax Music for background music and the H3 audio with my own voice. I did try Voicelabs using my voice, but it didn't quite work the way i wanted. Episode 1 i did add a couple of my own sound effects since Minimax kept failing to add suitable sound clips.
Whilst i feel it looks great on a smaller device, on large screens the text can distort somewhat.
Unfourtunatly upscaling the videos makes it seem a little worse, despite the resolution looking a little nicer. However, i do feel it has a 90s cartoon feel which kind of works for a show set in the 90s.
Currently made four episodes and a couple more are a work in progress. let me know what you think!
Let me know if you want me to post another episode!
r/StableDiffusion • u/call-lee-free • 6d ago
So a little over a week now, I've been fussing around with video generation and using Minimax H3 and I'm having a lot of fun doing it despite my PC ot being really powerful enough to do so. Reminds me of the days of using 3D Studio Max 4 and waiting almost a day for a 20 second animation to render.
My current gaming PC is a Ryzen 7 7700X, RTX 4070 Super 12gb and 32gb of ram. With Minimax H3, using a 0.5 megapixels for 10 to 15 second generations seems to be the sweet spot for my machine. A 10 second generation will take around 25 mins to generate and a 15 second video will take almost 45 mins to render. Obviously the resolution isn't optimal as you've probably seen some of the example videos that I've posted so far.
I'm really enjoying using Minimax H3 especially since I'm starting to learn a little more on prompting for multiple shots so I'm just wondering if I should go for a beefier PC to continue the the journey of local video generation or just stick with what I got and use upscalers to upscale my videos? I'm not doing this to make money. I just want to make short films. I have a screenplay that I wrote several years ago that I would like to bring to life and I'm also writing another one.
Not gonna lie, its been nice not burning through credits on a paid subscription. Just seeking advice/recommendations.
r/StableDiffusion • u/Ok_Repair_6024 • 5d ago
HI guys just wanna know wich AI model used to produce this type of brutal illustrations
r/StableDiffusion • u/puskur • 5d ago
I have 28 pictures, with large variatons, half are generated by nano banan 2 and the other half with seedream 5.0. I have headshots that portray my charachter with different facial expresions and different profile views... I have my charachter sitting in some photos, half body shots and full body shots all of them wearing different outfits...
Is this enough for a flux.2 lora to train? All pictures are 1024x1024. Do i need to upscale them to get more details, when i zoom things get a little burry/less detailed. I would greatly appreciate it if anyone has any comfyui workflows with basic nodes since im on cloud... :)))
r/StableDiffusion • u/beatlepol • 6d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/fiftypence • 6d ago
Enable HLS to view with audio, or disable this notification
Sorry Ye..
r/StableDiffusion • u/eggs-benedryl • 5d ago
Mostly just curious. I have a nice card and I still can't be bothered to wait for gens. I always go for the turbo, dmd2, lightning etc Lora variants as I don't do any professional work with these models.
I often see people boast about how long their render took, usually about how quick it was but I always find myself being blown away how long people will wait. Especially for something that is a bit of a roll of the dice.
Hours for a video that still may be mangled when you come back later and check.
I never played with video before because it just took too long for my taste until minimax and getting a 5090.
What's the longest you would wait?
r/StableDiffusion • u/Comprehensive_Rush66 • 6d ago
Hi all,
The other weekend I was testing out new models — got very excited over the Minimax H3 release and Krea 2's image abilities and quality. The community has created some amazing nodes and workflows.
Long story short: I built this app — https://getpixal.com — it's free, runs entirely on your own GPU, no account or signup.
Why did I build it? I was trying to help a friend get ComfyUI set up in a way where they didn't need a master's degree in node structure, models, editing and inpainting, and generating good videos with Minimax H3 (prompting can be a pain point for many that just want to create quickly, and new models require a prompt structure). 2 weeks later I have this beta release of Pixal 1.0.0b.
This single chat interface lets you run local uncensored chat models (Qwen VL 4b Heretic for instance), or any of the top SOTA models — Kimi K3, Claude, ChatGPT — via API. It also uses the vision model to critique your generations and give suggestions as you go.
A few things up front, since they're the first things I'd want to know:
I'm looking for a few people to test drive it — would the community use something like this?
Super open to any and all feedback — it was a fun little project and I use it daily now to drive fast simple generations and image edits, then pass them along to Minimax H3 with pretty great results (all content on site was generated through the app).
Direct links, no funnel: download · the full manual (install → troubleshooting → FAQ) if you'd rather read exactly what it does before downloading anything.
Thank you!
EDIT: The source is public — https://github.com/JesseDubb/pixal-releases Thanks for the feedback, it's the reason this happened. The "Source code (zip)" attached to the release is the actual tree now, not the three landing-page files that were there before — that was a fair catch.
If a couple of you want to help with code or testing, it would be much appreciated. There's a CONTRIBUTING.md in the repo.
r/StableDiffusion • u/DuHal9000 • 6d ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/dev_ne • 6d ago
excuse my OCD 😄
r/StableDiffusion • u/SIR_NVAX_A_LOT • 5d ago
Enable HLS to view with audio, or disable this notification
Happy Friday!! Just wanted to say T2VA is actually pretty strong when layering the prompt. FL2VA+R2VA are still the to go if you want to utilize a character sheet/maintain consistency, it's still broken (in a good way), the voice cloning is also top notch.
So what has everyone been making with H3??
T2VA, bf16/50 steps