Generate
Back to Models
Models/kwaivgi/video · text-to-video · image-to-videoPass-through

Kling Video 3.0

Kling Video 3.0 is designed for cinematic dynamics, directed camera movement, and audio-aware generation from text or a starting image. MachGen exposes text-to-video and image-to-video, including an optional end frame. For Reference to Video, use Kling Video 3.0 Omni.

Current MachGen support
Text to VideoImage to Video3-15s720p / 1080p / 4KOptional audio
Starting at $0.076 $0.067 / output s with auto-reload
Auto-reloadRegular
720p · audio off$0.067/s$0.076/s
1080p · audio off$0.09/s$0.101/s
2160p · audio off$0.336/s$0.378/s
720p · audio on$0.101/s$0.113/s
1080p · audio on$0.134/s$0.151/s
2160p · audio on$0.336/s$0.378/s
Live public pricebookSee full pricebook

Overview

Kling Video 3.0 generates short video from text or a starting image with optional native audio. Image-to-video can also use an optional end frame.

Reference-to-video is not exposed on this model. Use Kling Video 3.0 Omni for reference images, reusable elements, and multi-shot reference workflows.

Model guide

Why choose Kling Video 3.0

Key strengths

  • Cinematic motion and camera behavior with strong subject and visual continuity.
  • Text-to-video and image-to-video share the same 3-15 second output range from 720p through 4K.
  • Image-to-video accepts a starting image and an optional end frame.
  • Native audio can coordinate dialogue, ambience, and effects with the action.

Good fit for

  • Action, fashion, automotive, and camera-led cinematic shots.
  • Prompt-led scenes with optional synchronized dialogue, ambience, or effects.
  • Image-led motion and transitions with optional opening and closing frame anchors.
Supported on MachGen

Inputs and output settings

This table reflects the model routes and controls currently exposed by MachGen.

ModeRequired inputAvailable settings
Text to VideoText promptResolution: 720p, 1080p, 4K
Aspect ratio: 16:9, 9:16, 1:1
Duration: 3-15s
Frame rate: 30fps
Optional generated audio
Prompt limit: 2,500 characters
Image to VideoText prompt and source image; optional end frameResolution: 720p, 1080p, 4K
Duration: 3-15s
Frame rate: 30fps
Optional generated audio
Prompt limit: 2,500 characters
Workflow

Using Kling Video 3.0

  1. Choose a supported generation mode in Playground.
  2. Add the required prompt and source or reference assets.
  3. Select output settings from the controls shown for that mode.
  4. Generate, then inspect the returned asset before reusing the settings through the API.

Practical guidance

  • Describe the subject, action or composition, environment, and visual direction in a clear order.
  • For image animation, describe the intended motion and camera behavior instead of repeating what is already visible.
  • When audio is enabled, include any dialogue, ambience, or sound cues that matter to the scene.

Modes available on MachGen

  • Text to video - describe the scene, movement, and optional audio.
  • Image to video - use a starting image and optional end frame to anchor appearance and framing.
  • Reference to video - use Kling Video 3.0 Omni instead of Kling Video 3.0.

Prompting notes

  • Order prompts as setting, subject detail, motion, camera, then audio.
  • Use compatible start and end frames when the final composition is fixed, and describe the transition between them.
  • Use Kling Video 3.0 Omni when the request needs reference images, reusable elements, or multi-shot reference control.
  • Review speech, lip synchronization, and identity continuity. Appearance can drift across separate jobs, so longer pieces usually need editing.
Best results

How to get better Kling Video 3.0 results

  • State the camera move separately from subject motion so the two directions do not compete.
  • Use compatible start and end frames when the final composition is fixed, and describe the transition rather than restating both images.
  • Use Kling Video 3.0 Omni for Reference to Video, reusable elements, or multi-shot reference workflows.
Limits

Limits and availability

  • MachGen admits only the modes and settings listed above for this catalog model.
  • Source assets must finish uploading before a request can be submitted.
  • Generated results can vary between requests, including when the same prompt and settings are reused.
  • Kling Video 3.0 is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
Learn more

Learn more

For model background, technical details, and architecture, see the official Kling AI developer overview.

Try this model in the MachGen Playground.

Official documentation may describe capabilities or parameters that are not currently exposed by MachGen. Use the Inputs and output settings section above as the source of truth for this page.

Pricing

Kling Video 3.0 pricing

See Pricing for currently published MachGen rates. When available, the Playground estimate reflects the selected task and output settings.

On MachGen

Kling Video 3.0 on MachGen

Deployment: Pass-through. Kling Video 3.0 is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.

The Playground and API sections on this page list the inputs and output settings exposed for this model.