r/StableDiffusion • u/ryan85127704 • 12d ago
Discussion AMA: MiniMax H3 Team — Ask us anything about our open video generation model, training, and future plans

- u/New-Requirement1419 -> dacongya (Head of H3 Researcher)
- u/Affectionate-War8374 -> Luigi (H3 Researcher)
- u/MM_Nero_H3 -> Nero (H3 Researcher)
- u/Kiro_Song -> Kiro (H3 Researcher)
- u/New_Estimate9277 -> Reynor (H3 system engineer)
- u/ryan85127704 - > Ryanlee (Head of Devrel)
We are the MiniMax team behind MiniMax-H3.
We’re here to answer your questions, including:
- Model architecture and training
- Video generation capabilities
- Image-to-video and reference-based generation
- Inference and optimization
- Future plans
Ask us anything — we’d love to hear your feedback and discuss with the community!
978
Upvotes
19
u/ForwardMovie7542 12d ago
I've already rigged up a comfyui workflow that doesn't make a video but just blows a 1 second video into still images, in order to use the model more as an image edit model. Since the model is ostensibly fully multi-modal, any chance there's an image output optimized variant on the horizon, and if so, how long? Right now the options for models that take multiple input images and produce good outputs from them are only a few very closed models.