Generate
Back to Models
Models/xai/video · text-to-videoPass-through

Grok Imagine Video

Grok Imagine Video favors expressive interpretation and quick creative loops for general-purpose clips, social concepts, and early visual development. MachGen exposes text-to-video from 1 to 15 seconds at 480p or 720p across landscape, portrait, square, and 4:3 framing.

Current MachGen support
Text to Video1-15s480p / 720p
Starting at $0.05 / output s
480p$0.05/s
720p$0.07/s
Live public pricebookSee full pricebook
Overview

Grok Imagine Video

Grok Imagine Video turns a scene description into short general-purpose clips for drafts, social concepts, and rapid exploration. It favors expressive interpretation and fast loops over ultra-rigid cinematic tooling.

Use it when you want a quick text-to-video read of an idea before committing to a heavier cinematic model.

Model guide

Why choose Grok Imagine Video

Key strengths

  • Fast, expressive text-to-video generation for ideas that benefit from visual personality.
  • Flexible duration and framing make it convenient for landscape, portrait, square, and 4:3 concepts.
  • A good exploration model when direction matters more than strict reference matching.

Good fit for

  • Social concepts, visual jokes, mood pieces, and early story exploration.
  • Quick motion studies before committing to a reference-heavy workflow.
  • Expressive scenes that can tolerate creative interpretation.
Supported on MachGen

Inputs and output settings

This table reflects the model routes and controls currently exposed by MachGen.

ModeRequired inputAvailable settings
Text to VideoText promptResolution: 480p, 720p
Aspect ratio: 16:9, 9:16, 1:1, 4:3
Duration: 1-15s
Frame rate: 24fps
Workflow

Using Grok Imagine Video

  1. Choose a supported generation mode in Playground.
  2. Add the required prompt and source or reference assets.
  3. Select output settings from the controls shown for that mode.
  4. Generate, then inspect the returned asset before reusing the settings through the API.

Practical guidance

  • Describe the subject, action or composition, environment, and visual direction in a clear order.
Capabilities

Modes

  • Text to video - describe the scene, motion, and mood; get a short clip back.
  • Exploration first - optimized for ideation speed rather than maximum directorial control.
Prompting

Prompting notes

  • One subject, one action, one camera idea per clip.
  • Lean into vivid verbs and atmosphere; add lens language when you need more control.
  • Keep duration short while exploring, then lengthen once the visual direction is established.
Best results

How to get better Grok Imagine Video results

  • Lead with one memorable subject and action, then add setting, mood, and camera direction.
  • Use concrete visual language instead of abstract adjectives such as impressive or beautiful.
  • Match the aspect ratio to the final channel before refining the prompt.
  • If the result becomes chaotic, shorten the duration or reduce the number of moving subjects.
Limits

Limits and availability

  • MachGen admits only the modes and settings listed above for this catalog model.
  • Source assets must finish uploading before a request can be submitted.
  • Generated results can vary between requests, including when the same prompt and settings are reused.
  • Grok Imagine Video is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
Learn more

Learn more

For model background, technical details, and architecture, see the official xAI video generation guide.

Try this model in the MachGen Playground.

Official documentation may describe capabilities or parameters that are not currently exposed by MachGen. Use the Inputs and output settings section above as the source of truth for this page.

Pricing

Grok Imagine Video pricing

See Pricing for currently published MachGen rates. When available, the Playground estimate reflects the selected task and output settings.

On MachGen

Grok Imagine Video on MachGen

Deployment: Pass-through. Grok Imagine Video is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.

The Playground and API sections on this page list the inputs and output settings exposed for this model.