Seedance 2.0 is ByteDance's high-quality multimodal video model for directed camera work, subject consistency, lip-synced dialogue, and native sound. On MachGen it supports text, image, mixed-reference, and first-to-last-frame generation; reference mode accepts up to nine images, three videos, and three audio clips, with 4-15 second output from 480p through 4K.
| Auto-reload | Regular | |
|---|---|---|
| 480p | $0.06/s | $0.07/s |
| 720p | $0.12/s | $0.14/s |
| 1080p | $0.30/s | $0.37/s |
| 2160p | $0.59/s | $0.72/s |
| 480p · with input video | $0.035/s | $0.04/s |
| 720p · with input video | $0.07/s | $0.085/s |
Seedance 2.0 is ByteDance's multimodal video model: text, stills, video clips, and audio share one input space, and a single generation returns picture plus soundtrack together. Dialogue, Foley, and score stay locked to the frames without a separate audio stage.
That shared input space lets you steer with references instead of only prose - a still for appearance, a clip for motion language, a track for rhythm - then describe how they should combine in plain English.
This table reflects the model routes and controls currently exposed by MachGen.
| Mode | Required input | Available settings |
|---|---|---|
| Text to Video | Text prompt | Resolution: 480p, 720p, 1080p, 4K Aspect ratio: Adaptive, 21:9, 16:9, 9:16, 1:1, 4:3, 3:4 Duration: 4-15s, Smart Frame rate: 24fps Optional generated audio Seed control Standard or high bitrate |
| Image to Video | Text prompt and source image; optional end frame | Resolution: 480p, 720p, 1080p, 4K Aspect ratio: Adaptive, 21:9, 16:9, 9:16, 1:1, 4:3, 3:4 Duration: 4-15s, Smart Frame rate: 24fps Optional generated audio Seed control Standard or high bitrate |
| Reference to Video | Per-type limits: up to 9 images, up to 3 videos, up to 3 audio files, plus a text prompt | Resolution: 480p, 720p, 1080p, 4K Aspect ratio: Adaptive, 21:9, 16:9, 9:16, 1:1, 4:3, 3:4 Duration: 4-15s, Smart Frame rate: 24fps Optional generated audio Seed control Standard or high bitrate |
| First & Last Frame | Text prompt, start frame, and end frame | Resolution: 480p, 720p, 1080p, 4K Aspect ratio: Adaptive, 21:9, 16:9, 9:16, 1:1, 4:3, 3:4 Duration: 4-15s, Smart Frame rate: 24fps Optional generated audio Seed control Standard or high bitrate |
@name token when the prompt should call it out.For model background, technical details, and architecture, see the official Seedance 2.0 page.
Try this model in the MachGen Playground.
Official documentation may describe capabilities or parameters that are not currently exposed by MachGen. Use the Inputs and output settings section above as the source of truth for this page.
See Pricing for currently published MachGen rates. When available, the Playground estimate reflects the selected task and output settings.
Deployment: Pass-through. Seedance 2.0 is pass-through. MachGen sends the request through its upstream provider integration and returns the generated asset.
The Playground and API sections on this page list the inputs and output settings exposed for this model.