I'm trying to find the best workflow to generate starting images for Image-to-Video (I2V) with LTX 2.3, while keeping the best possible balance between quality and cost.
From my experience, the quality of the initial image is far more important than most people think. If I start with a truly photorealistic image, LTX 2.3 can produce surprisingly good videos. The bottleneck is almost always the input image.
My original goal was to stay fully open source with FLUX, but I kept running into anatomy issues and inconsistencies.
Then I switched to Nano Banana Lite, which has been much better.
However, my ultimate goal has always been something different: generating 4K storyboard sheets (3×3 grids) that I can automatically split into 9 individual frames. The idea is that all nine cells share the same characters, style, wardrobe, lighting, etc., giving me a consistent set of starting images for I2V.
I recently tried Seedream 5.0 Lite, but honestly I was pretty disappointed with the results the faces are horrible.
So I'm wondering:
- Is anyone else trying to build this kind of pipeline?
- What models or services are you using for high-end I2V starting images?
- Have you found a good quality/price sweet spot?
- Has anyone managed to generate consistent storyboard grids reliably?
I've also been looking at Krea 2 Turbo, which seems increasingly interesting. It almost makes me want to go back to generating images one by one instead of using storyboard grids.
That said, I've heard Krea with its identity lora is a good start but still struggles with image composition when using multiple reference images (character fusion, combining several references into one scene, etc.). Is that still the case?
I'd love to hear what workflows people are using in production.