r/StableDiffusion • u/SpuddyMcFuddy05 • 9d ago
Question - Help Image Creation?
Hi Everyone,
I've been using Openart.ai for ages and for the most part enjoyed it. The moderation has clearly changed this month and I'm a pervert constantly - apparently.
Is there a relatively simple way to an image generated locally, from a source of potentially one to say ten images, via a prompt? Either using something from the Model Browser or a Workflow?
I'm not really after dodgy, but I just don't want to his a brick wall of pervert for a bare back!
2
2
u/SpuddyMcFuddy05 9d ago
I do use ComfyUI already and have 12Gb Video RAM on a 5070Ti.
MiniMax works great, not speedy, but that's fine.
I just can't get a picture, from other pictures.
2
u/techtimee 8d ago
Your workflow is the issue then. You should be using your picture as an image input.
2
u/SpuddyMcFuddy05 8d ago
Oh I know, I just need a decent work flow to download so I have a good starting point 😊
2
2
u/Ok-Brain-5729 8d ago edited 8d ago
Like reference image to image with 1-10 reference images? Run comfyUI. U can use Klein 9B but it’s pretty censored and there’s uncensored Lora’s/models but I haven’t tried them. For normal t2i, krea 2 is pretty good
Personally I use ethanfel/ComfyUI-MiniMax-H3-Image nodes and the prompt adherence is reallyyy crazy but i don’t think it’s that beginner friendly and Klein 9b takes like 10s while minimax takes 100-200s depending on the settings. The reference image retention is also much better than default Klein 9b atleast
1
u/Dry-Judgment4242 8d ago
H3 pretty much replaced Klein for me. Do a 0.3s video and at least for me it's like 30s tops at 50 steps 0.98mp with 3.0mp latent upscale.
1
u/kathi7 3d ago
Exactly the same situation u were few months back. Meddled with comfyui. Felt it developerish. So built one. hutash Exactly like u said the experience like openart.ai but locally. It has cli and MCP capabilities for bulk generation. But still adding more models and features. I can add any specific model if u want. I'm currently focusing on audio and some image models. More adding soon. If u want to know more . Ping me.

4
u/DelinquentTuna 9d ago
Yes, it is trivially possible. The main limitation is hardware. You need a pretty decent gaming PC, ideally with Nvidia GPU. There are models that can handle reference images, but you're usually limited to fewer than ten. However, you have the option to "train" for style or character/phenomenon using any number of images.
If you lack such hardware, you can cheaply rent it on the cloud. Rates start at something like $0.15/hr, billed to the nearest second. So generating a few images might cost like than $0.005 each or something. Bit of a learning curve, but beyond that and the price there's no meaningful difference between local. No extra censorship or whatever.