Overview
Grok Imagine Video
Grok Imagine Video turns a scene description into short general-purpose clips for drafts, social concepts, and rapid exploration. It favors expressive interpretation and fast loops over ultra-rigid cinematic tooling.
Use it when you want a quick text-to-video read of an idea before committing to a heavier cinematic model.
Model guideWhy choose Grok Imagine Video
Key strengths
- Fast, expressive text-to-video generation for ideas that benefit from visual personality.
- Flexible duration and framing make it convenient for landscape, portrait, square, and 4:3 concepts.
- A good exploration model when direction matters more than strict reference matching.
Good fit for
- Social concepts, visual jokes, mood pieces, and early story exploration.
- Quick motion studies before committing to a reference-heavy workflow.
- Expressive scenes that can tolerate creative interpretation.
Supported on MachGenInputs and output settings
This table reflects the model routes and controls currently exposed by MachGen.
| Mode | Required input | Available settings |
|---|
| Text to Video | Text prompt | Resolution: 480p, 720p Aspect ratio: 16:9, 9:16, 1:1, 4:3 Duration: 1-15s Frame rate: 24fps |
WorkflowUsing Grok Imagine Video
- Choose a supported generation mode in Playground.
- Add the required prompt and source or reference assets.
- Select output settings from the controls shown for that mode.
- Generate, then inspect the returned asset before reusing the settings through the API.
Practical guidance
- Describe the subject, action or composition, environment, and visual direction in a clear order.
Capabilities
Modes
- Text to video - describe the scene, motion, and mood; get a short clip back.
- Exploration first - optimized for ideation speed rather than maximum directorial control.
Prompting
Prompting notes
- One subject, one action, one camera idea per clip.
- Lean into vivid verbs and atmosphere; add lens language when you need more control.
- Keep duration short while exploring, then lengthen once the visual direction is established.
Best resultsHow to get better Grok Imagine Video results
- Lead with one memorable subject and action, then add setting, mood, and camera direction.
- Use concrete visual language instead of abstract adjectives such as impressive or beautiful.
- Match the aspect ratio to the final channel before refining the prompt.
- If the result becomes chaotic, shorten the duration or reduce the number of moving subjects.
LimitsLimits and availability
- MachGen admits only the modes and settings listed above for this catalog model.
- Source assets must finish uploading before a request can be submitted.
- Generated results can vary between requests, including when the same prompt and settings are reused.
- Grok Imagine Video is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
Learn moreLearn more
For model background, technical details, and architecture, see the official xAI video generation guide.
Try this model in the MachGen Playground.
Official documentation may describe capabilities or parameters that are not currently exposed by MachGen. Use the Inputs and output settings section above as the source of truth for this page.
PricingGrok Imagine Video pricing
See Pricing for currently published MachGen rates. When available, the Playground estimate reflects the selected task and output settings.
On MachGenGrok Imagine Video on MachGen
Deployment: Pass-through. Grok Imagine Video is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
The Playground and API sections on this page list the inputs and output settings exposed for this model.