Overview
Grok Imagine Video 1.5 turns still images into expressive video with realistic motion, object interaction, and automatically generated audio matched to the scene. It works well for product animation, portraits, illustrations, and keyframes that need motion without rebuilding the composition from text.
MachGen exposes image-to-video only. Upload one still, describe the action, camera movement, atmosphere, and sound, then generate a 1-15 second clip. Use the separate Grok Imagine Video catalog entry when you need text-to-video.
Model guideWhy choose Grok Imagine Video 1.5
Key strengths
- Image-to-video generation centered on realistic motion and interaction from an existing frame.
- Automatically synchronized music, ambience, or effects can make a still feel like a complete scene.
- The source image provides strong control over composition, subject, and visual style.
Good fit for
- Animating portraits, product stills, illustrations, and key art.
- Adding subtle environmental motion and sound to an approved frame.
- Short social assets that begin from a recognizable image.
Supported on MachGenInputs and output settings
This table reflects the model routes and controls currently exposed by MachGen.
| Mode | Required input | Available settings |
|---|
| Image to Video | Text prompt and source image | Resolution: 480p, 720p, 1080p Duration: 1-15s Frame rate: 24fps |
WorkflowUsing Grok Imagine Video 1.5
- Choose a supported generation mode in Playground.
- Add the required prompt and source or reference assets.
- Select output settings from the controls shown for that mode.
- Generate, then inspect the returned asset before reusing the settings through the API.
Practical guidance
- For image animation, describe the intended motion and camera behavior instead of repeating what is already visible.
Prompting notes
- Describe the change rather than repeating everything already visible in the source image.
- Use specific motion verbs and name the camera move when framing should change.
- Keep the sequence readable: subject action first, camera second, then atmosphere and sound.
- Describe desired audio directly in the prompt, including ambience, effects, music, or a short spoken line.
- Prefer one primary subject and action per clip; add complexity in later iterations.
Best resultsHow to get better Grok Imagine Video 1.5 results
- Use a sharp source image with a believable pose and enough room for the requested motion.
- Describe the movement trajectory and speed, including what should remain still.
- Mention the desired ambience or effects when sound is important to the scene.
- Avoid asking for a major camera move and a large subject transformation in the same short clip.
LimitsLimits and availability
- MachGen admits only the modes and settings listed above for this catalog model.
- Source assets must finish uploading before a request can be submitted.
- Generated results can vary between requests, including when the same prompt and settings are reused.
- Grok Imagine Video 1.5 is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
Learn moreLearn more
For model background, technical details, and architecture, see the official xAI video generation guide.
Try this model in the MachGen Playground.
Official documentation may describe capabilities or parameters that are not currently exposed by MachGen. Use the Inputs and output settings section above as the source of truth for this page.
PricingGrok Imagine Video 1.5 pricing
See Pricing for currently published MachGen rates. When available, the Playground estimate reflects the selected task and output settings.
On MachGenGrok Imagine Video 1.5 on MachGen
Deployment: Pass-through. Grok Imagine Video 1.5 is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
The Playground and API sections on this page list the inputs and output settings exposed for this model.