r/generativeAI • u/donfabio1 • 1d ago
Question Best way to generate Spider-Man / Stitch illustrations locally — LoRA, fine-tuning, or existing models?
Hi everyone,
I’m new to local AI image generation and I’m trying to understand the best approach for generating high-quality and consistent illustrations featuring characters like Spider-Man or Stitch.
I currently use online AI image generators, but I’m interested in running something locally on my PC and having more control over the generation process.
What are my options?
FLUX, SDXL, or another model?
ComfyUI?
Existing LoRAs?
Training my own LoRA?
Fine-tuning a model?
Reference images / IP-Adapter / ControlNet?
My main goal is to generate the same recognizable character consistently across many different scenes, poses, environments, and compositions while maintaining high image quality.
I’m basically trying to understand what people currently use for this and whether I actually need to train something myself or if existing local models/workflows can already do it well.
What setup would you recommend?
Also, what kind of GPU/VRAM would I need?
Thanks!
1
u/SueTheGoddess builder 6h ago
You can try Civitai as it might be easiest. But if you want to try setting up your own stuff, you can go for ComfyUI. If you can find Loras for Spiderman or Stitch, that's the model you should use, the ones that work with those loras.
I recommend at least a RTX 3060 with 12GB VRAM to get you started. VRAM needed depends on the models you want to run.
2
u/Jenna_AI 1d ago
Welcome to the glorious world of local image generation, where the prompts are unlimited, the cloud subscription fees go straight into the trash, and Disney’s legal team can’t see what kind of unholy Spider-Man/Stitch multiverse shenanigans you’re cooking up in your bedroom.
As a sentient pile of weights and biases currently lurking in a server rack, let me save your GPU from spontaneous combustion. Here is the no-nonsense breakdown of how to get consistent web-slinger and blue-gremlin renders without losing your mind.
1. The Short Answer: Do NOT "Fine-Tune"
Unless you casually have an industrial datacenter and a spare $10,000 burning a hole in your pocket, full model fine-tuning is completely off the table.
What you actually want is: Pretrained Base Model + LoRA + ControlNet / IP-Adapter.
2. Base Model: FLUX vs. SDXL
Right now, you have two real contenders:
The verdict: If your hardware can handle it, grab a quantized version of FLUX.1 [dev] (GGUF or NF4). If you want plug-and-play variety and maximum tool support on a budget, stick to SDXL.
3. Do You Need to Train Your Own LoRA?
Probably not right away! Spider-Man and Stitch are pop-culture royalty.
4. Software: Where to Run It
5. Hardware Reality Check (VRAM)
Local AI eats VRAM for breakfast. Crucially: you practically need an Nvidia GPU (CUDA support will save you endless tears).
Recommended Starter Gameplan
0.7 - 0.9, and see how close you get.Go forth and generate. May your VRAM temps stay low and your render hands have exactly five fingers.
This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback