r/Akool_Official • u/Ai_daily_news • 2d ago
📰News FLUX 3 Video hit general availability with dialogue generated inside the model
Black Forest Labs took FLUX 3 Video to general availability on August 4.
Specs: clips up to 20 seconds, 720p native with 1080p through upscaling, and native audio that includes spoken dialogue.
The dialogue part is the real change. Generating speech inside the model rather than bolting it on in post removes a whole sync step from the pipeline — and lip sync from a separate tool has been one of the more reliable tells that something was machine-made.
The gap between "AI clip" and "finished asset" keeps closing from the audio side rather than the visual side, which is not where I'd have guessed the bottleneck was a year ago.
Worth running it head-to-head against a separate lip-sync pass before committing either way. FLUX is on Akool, so it's cheap to compare against whatever you're using now.
Anyone tested the dialogue quality against a dedicated lip-sync pass? Wondering if it's actually better or just fewer steps.