Generate

Products / Models

Models

Browse video and image models on MachGen - including Text to Video, Image to Video, Reference to Video, Video Upscaling, Text to Image, and Image Editing.

Showing 37 of 39 models

T2VI2VR2VF2F

Minimax / minimax-h3

MiniMax H3

MiniMax H3 for text, image, frame-pair, and multimodal reference video with native stereo audio.

Starting at $0.04 / output s
View model ->
T2VI2VR2VF2F

ByteDance / seedance-2.5

Seedance 2.5

Seedance 2.5 for 30-second native video with up to 50 multimodal references and directed editing.

Starting at $0.085 / output s
View model ->
T2VI2VF2F

Lightricks / ltx-2.3-pro

LTX 2.3 Pro

LTX 2.3 Pro for cinematic 4K video, directorial control, and synchronized native audio.

Starting at $0.007 / output s
View model ->
T2VI2VR2VF2F

Vidu / q3-turbo

Vidu Q3 Turbo

Vidu Q3 Turbo for fast, balanced video iteration across text, image, reference, and frame-pair workflows.

Starting at $0.015 / output s
View model ->
T2VI2V

Wan / wan-2.2-i2v-a14b

Wan 2.2 A14B

Wan 2.2 A14B for rich visuals and detailed motion from text or stills, served on low-cost managed inference.

Starting at $0.018 / output s
View model ->
T2VI2VR2VF2F

ByteDance / seedance-2.0

Seedance 2.0

Seedance 2.0 for cinematic multimodal video, coherent multi-shot storytelling, and synced native audio.

Starting at $0.06 / output s
View model ->
T2VR2V

Kuaishou / kling-o3

Kling Video 3.0 Omni

Compose video from prompts, reference images, and Kling elements

Starting at $0.067 / output s
View model ->
T2VI2V

Kuaishou / kling-v3-video

Kling Video 3.0

Kling Video 3.0 for cinematic text-to-video and image-to-video with optional native audio.

Starting at $0.067 / output s
View model ->
T2VI2VR2VF2F

Alibaba / wan-3.0

Wan 3.0

Wan 3.0 for production-scale video generation with 30-second shots, multimodal direction, and synchronized sound.

Starting at $0.06 / output s
View model ->
R2V

Vidu / q3-r2v

Vidu Q3

Vidu Q3 for high-quality reference-to-video with look-locked subjects, prompt-directed camera motion, and optional audio.

Starting at $0.035 / output s
View model ->
T2VI2VF2F

Vidu / q3

Vidu Q3 Pro

Vidu Q3 Pro for high-fidelity text, image, and frame-pair video with optional synchronized audio.

Starting at $0.045 / output s
View model ->
I2V

Vidu / q3-pro-fast

Vidu Q3 Pro Fast

Vidu Q3 Pro Fast for low-latency, high-quality image animation with prompt-directed camera motion and optional audio.

Starting at $0.10 / output s
View model ->
T2V

Google / veo-3.1

Veo 3.1 Standard

Google Veo 3.1 Standard for high-quality cinematic text-to-video with synchronized native audio.

Starting at $0.40 / output s
View model ->
T2V

Google / veo-3.1-fast

Veo 3.1 Fast

Google Veo 3.1 Fast for cinematic native-audio video with quicker, lower-cost prompt iteration.

Starting at $0.10 / output s
View model ->
T2VI2VR2V

PixVerse / pixverse-v6

Pixverse V6

PixVerse V6 for cinematic, stylized video with references, prompt-directed camera motion, and optional audio.

Starting at $0.02 / output s
View model ->
T2VI2VR2V

PixVerse / pixverse-c1

Pixverse C1

PixVerse C1 for cinematic action, VFX-heavy scenes, and reference-guided visual continuity.

Starting at $0.024 / output s
View model ->
T2VI2VR2V

Alibaba / happyhorse-1.1

HappyHorse 1.1

HappyHorse 1.1 for high-quality general video and stronger multi-reference subject consistency.

Starting at $0.112 / output s
View model ->
T2VI2VR2V

Alibaba / happyhorse-1.0

HappyHorse 1.0

HappyHorse 1.0 for balanced-cost, general-purpose video from text, a first frame, or references.

Starting at $0.112 / output s
View model ->
T2V

xAI / grok-imagine-video

Grok Imagine Video

Grok Imagine Video for fast, expressive text-to-video concepts and social-ready visual exploration.

Starting at $0.05 / output s
View model ->
I2V

xAI / grok-imagine-video-1.5

Grok Imagine Video 1.5

Grok Imagine Video 1.5 for realistic still-image motion, object interaction, and synchronized generated audio.

Starting at $0.08 / output s
View model ->
UPSCALE

Topaz / video-precision

Topaz Precision

Topaz Precision for detail-preserving video upscaling to 1080p, 2K, 4K, or 8K.

Starting at $0.127 / credit
View model ->
UPSCALE

Topaz / video-precision-animation

Topaz Precision Animation

Topaz Precision Animation for upscaling hand-drawn and cel animation.

Starting at $0.127 / credit
View model ->
UPSCALE

Topaz / video-generative

Topaz Generative

Topaz Generative for diffusion-based video restoration up to 4K.

Starting at $0.127 / credit
View model ->
UPSCALE

Topaz / video-generative-fast

Topaz Generative Fast

Topaz Generative Fast for cheaper generative video restoration on longer clips.

Starting at $0.127 / credit
View model ->
T2II2I

Black Forest Labs / flux-2-dev

FLUX.2 Dev

FLUX.2 Dev for high-quality open-weight image generation, multi-image editing, and fast experimentation.

Starting at $0.004 / MP
View model ->
T2I

Hidream / hidream-o1

HiDream O1

HiDream O1 for high-detail text-to-image across fantasy, illustration, and photoreal visual styles.

Starting at $0.003 / MP
View model ->
T2I

xAI / grok-imagine-image

Grok Imagine Image

Grok Imagine Image for fast, low-cost text-to-image with a bold, expressive visual style.

Starting at $0.02 / image
View model ->
T2II2I

xAI / grok-imagine-image-quality

Grok Imagine Image Quality

Grok Imagine Image Quality for higher-fidelity generation, cleaner detail, and reference-based editing.

Starting at $0.05 / image
View model ->
T2II2I

Google / nano-banana-2

Nano Banana 2

Nano Banana 2 for fast, high-fidelity Gemini image generation, natural-language edits, and clear typography.

Starting at $0.034 / image
View model ->
T2II2I

Google / nano-banana-pro

Nano Banana Pro

Nano Banana Pro for high-quality Gemini image generation, complex briefs, typography, and controlled edits.

Starting at $0.067 / image
View model ->
T2II2I

ByteDance / seedream-5.0-lite

Seedream 5.0 Lite

Seedream 5.0 Lite for efficient high-resolution image creation, layout-aware prompts, and reference edits.

Starting at $0.032 / image
View model ->
T2II2I

OpenAI / gpt-image-2

GPT Image 2

GPT Image 2 for high-fidelity image generation, precise instruction following, typography, and revisions.

Starting at $0.04 / image
View model ->
IMAGE_UPSCALE

Topaz / image-precision

Topaz Image Precision

Topaz Image Precision for faithful 2x or 4x image upscaling.

Starting at $0.127 / credit
View model ->
IMAGE_UPSCALE

Topaz / image-generative

Topaz Image Generative

Topaz Image Generative for prompt-steerable image upscaling.

Starting at $0.127 / credit
View model ->
T2ST2D

Elevenlabs / eleven-v3

ElevenLabs v3

ElevenLabs v3 for expressive speech and multi-speaker dialogue generated in one pass.

Starting at $0.0001 / char
View model ->
T2SFX

Elevenlabs / eleven-sfx-v2

ElevenLabs Sound Effects v2

ElevenLabs SFX v2 for foley and designed sound effects written as a description.

Starting at $0.002 / output s
View model ->
T2M

Elevenlabs / eleven-music-v2

ElevenLabs Music v2

ElevenLabs Music v2 for instrumental and vocal tracks, from a prompt or a section plan.

Starting at $0.005 / output s
View model ->