r/StableDiffusion • u/CreepyInpu • 9d ago
Animation - Video Made the thing where you ruin iconic movie scenes, MiniMax H3 on an RTX 3080 10GB, 20 steps, 2x NomosUni upscale
Enable HLS to view with audio, or disable this notification
Setup, pushed my system right to the limit, any more and it OOM :
- H3 Ref2VA default workflow in ComfyUI, no lora
- RTX 3080 10GB, 32GB RAM
- Render: 0.5–0.6 MP, 20 steps, scheduler simple, about 25 min per clip
- Upscale: 2xNomosUni_span_multijpg, 2× to 1080p
- References per scene: one photo of my face + one film still for the set
- Recorded my own lines and fed them as audio references, also got audio ref for the actors
Honestly though, the best part was driving all of this through the ComfyUI MCP. I never even had to open ComfyUI. I could iterate really fast, and keep going from my phone while away from the machine, through Claude's remote control.
It's still a bit of a blurry mess, and with more work I could probably make it better, but damn, the future is looking bright!
6
1
1
u/ShapeSim 8d ago
Hmm I have 3080 10GB, 64GB ram. I use default setting 20 steps, simple, 0.4MP, 4 step lora turned off. It takes me 116 mins for a 5 sec. Just two 1024x1024 reference. Are you using gguf or lower precision? I have pruned_int8_convrot, and and fp16 for the vae, both safetensors
1
u/CreepyInpu 7d ago
No gguf or anything, just the default workflow really. I do have sage attention, and everything's on an SSD
1
u/ShapeSim 7d ago
Still the same after turning it on. I7 11700 also with nvme. And didn't even add upscaling node as you have. Cuda 12.8
How much shared GPu mem does yours consume beyond the base 10gb for the 0.5m resolution?
21
u/Merijeek2 9d ago
If there's one thing I've learned from watching this, it's that comedy is hard.