Models / Grok Imagine / Text to video
Grok Imagine text to video prompting guide
Last checked 2026-10-11Grok Imagine Video 1.5 (xAI) · grok-imagine-video-1.5, 1.5 Lite, grok-imagine-video (edit/extend)
How to prompt Grok Imagine in text-to-video mode. Every point links its source; community tips are marked. Your agent gets the same guide through the Atlas MCP (get_guide).
Prompt structure
opening frame→subject→action→camera→pacing→audio
- No dedicated prompting guide published; describe subject, camera, pacing and sound in plain language. source ↗
What the text should describe
- t2v first generates a frame from the prompt then animates it: describe the opening frame clearly, then the motion, camera, pacing and sound. source ↗
Responds well to
- t2v first generates a frame from your prompt then animates it, so the opening frame description matters. source ↗
Duration and limits
- Duration 1–15 s (default 8). 1.5 native 1080p for t2v/i2v; r2v up to 720p. i2v follows the image's aspect ratio. Editing keeps source length (max 8.7 s), 720p. source ↗
Negative prompt
Supported via .
Next
All Grok Imagine modesGrok Imagine model, prices & shotsFrom $0.02/s · pricesSeedance text to videoKling text to videoVeo text to videoRunway text to video