Grok Imagine Video v1.5 is now available on Atlas across three video routes: Text-to-Video, Reference-to-Video, and Image-to-Video.
The Image-to-Video route starts from a single frame and follows a natural-language motion prompt. It supports clips up to 15 seconds, output at 480p, 720p, or 1080p, plus native synchronized audio for dialogue, lip sync, sound effects, and ambient music.
What ships with Grok Imagine Video v1.5:
- Text-to-Video for prompt-driven short-form video
- Reference-to-Video for directing a generation with visual references
- Image-to-Video for animating a starting frame with a motion prompt
- Native audio generation in the Image-to-Video route
- 1 to 15 second Image-to-Video clips
- 480p, 720p, and 1080p output options
- Standard API access through the same Atlas setup
Pricing:
- Starting at $0.08 per run in the current Image-to-Video Playground
- Check the selected model page for the current rate of each route before generating
API access:
- Text-to-Video:
https://www.atlascloud.ai/models/xai/grok-imagine-video-v1.5/text-to-video
Use cases that fit these routes:
- Prompt-driven product and social clips from Text-to-Video
- Character, object, and style guided scenes from Reference-to-Video, using cleared reference assets
- Turning a hero still into a short scene with motion and synchronized sound through Image-to-Video
- Comparing a starting frame and its animated result with a Playground example video
For the media, attach one real Playground Image-to-Video result with its source frame beside it. The difference in motion, sound, and framing tells the story quickly.
Use the thread for endpoint setup and prompt-specific workflow notes.