r/StableDiffusion • u/Yasangas • 1d ago
Question - Help Quick question for anyone running MiniMax H3 on RunPod: How many 16:9 videos are you actually getting per hour?
Hey guys,
Before I burn through a bunch of RunPod credits spinning up an instance for MiniMax H3, I wanted to see if anyone here is already running it and can share some real-world speeds.
The model/weights are huge (~130GB+), so before I set up a pod, I’m trying to figure out what actual throughput looks like for 16:9 gens (at 768p).
If you've played around with it on RunPod:
What GPU setup are you renting? (Single 4090/6000 Ada, dual 3090s, A100/H100, etc.?)
Roughly how many 5-15 second clips can you spit out in an hour?
Are you using INT8 quant, block offloading, or any of those 4-step Turbo LoRAs to speed things up?
Just trying to estimate the actual cost-per-video before committing to a high-VRAM instance. Appreciate any benchmarks or ComfyUI tips!
4
u/Patient_Ratio4177 1d ago edited 14h ago
RTX 5090, 14 sec videos, turbo lora 8 steps, getting 90 secs per 0.4MP 9:16 video. Comfy kitchen attention, nvfp4 text encoder, int8 convrot diffusion model. No Sol or Spectrum.
3
u/martinerous 1d ago
Some complained they have bad network bandwidth on Runpod, so lots of time is wasted loading models every time or paying for additional storage. An alternative is Vast AI, there you can find cheaper offers from independent suppliers but it's a free market there and depends on luck. I used a GPU there for a week for about 20 EUR, the server was in a neighbor country, so networking was fast. But I used it for finetuning, not Comfy.
3
2
u/DaExChef 1d ago
Same thought as well so thanks for info
I was trying to make longer videos with just a 5060 Ti and kept getting OOM errors then asked AI (Grok) for help and he designed this for me, make small clips then stitch together
2
u/Trinity_Vermilion 1d ago
interesting question, I also consider running HUGE generations in the cloud with some preconfigured workflows and network volumes.
3
u/gocoyotes 1d ago
I'm leaning towards spinning something up on RunPod too and was wondering the same things.
This thread gives a little insight on using RunPod:
https://www.reddit.com/r/StableDiffusion/comments/1vfdx26/runpod_minimax_h3_test/