THE MODEL CATALOG

Video model APIs & pricing

Find text-to-video, image-to-video and video processing APIs by task. Compare Wan versions, inspect each model’s parameters and submit asynchronous jobs with one Jevrouter key.

532 models · Page 3 of 12

bytedance

seedance v1.5 Pro · image to video

Seedance 1.5 Pro (Image-to-Video) generates cinematic, live-action–leaning clips from a text prompt plus a first-frame image, preserving the image’s subject and composition while adding expressive motion and stable aesthetics. It supports 4–12s duration control (including Smart Duration), adaptive aspect ratio that follows the input image, and reproducible outputs via seeds—ideal for ad creatives and short-drama shots that need a strong visual anchor.

Video
Per video · quoted before submission
bytedance

seedance v1.5 Pro · text to video fast

Seedance 1.5 Pro Fast (Text-to-Video) generates cinematic, live-action–leaning clips from text with strong prompt adherence, expressive motion, and stable aesthetics. It supports 4–12s duration control, multiple aspect ratios, and reproducible generation via seeds—ideal for ads and short-drama workflows.

Video
Per video · quoted before submission
bytedance

seedance v1.5 Pro · text to video

Seedance 1.5 Pro (Text-to-Video) generates cinematic, live-action–leaning clips from text with strong prompt adherence, expressive motion, and stable aesthetics. It supports 4–12s duration control (including Smart Duration), multiple aspect ratios (including adaptive), and reproducible generation via seeds—ideal for ads and short-drama workflows.

Video
Per video · quoted before submission
bytedance

seedance v1.5 Pro · video extend fast

Seedance 1.5 Pro Fast Video-Extend turns short video clips into longer videos with natural motion continuation, stable aesthetics, and upscaled output. It supports 4-12s duration control, 720p/1080p resolutions, and reproducible generation via seeds—ideal for extending ad creatives and short-drama shots with higher quality output.

Video
Per video · quoted before submission
bytedance

seedance v1.5 Pro · video extend

Seedance 1.5 Pro Video-Extend turns short video clips into longer videos with natural motion continuation and stable aesthetics. It supports 4-12s duration control, multiple resolutions, and reproducible generation via seeds—ideal for extending ad creatives and short-drama shots.

Video
Per video · quoted before submission
bytedance

seedance v1 lite Image to Video 1080p

Seedance Lite creates coherent multi-shot image-to-video clips in 1080p with smooth, stable motion and accurate prompt following. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 lite Image to Video 480p

ByteDance Seedance V1 Lite is a 480p Image-to-Video model for coherent multi-shot videos with smooth motion and precise prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 lite Image to Video 720p

ByteDance Seedance Lite i2v 720p creates coherent multi-shot image-to-video clips with smooth, stable motion and prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 lite Text to Video 1080p

Seedance v1 Lite is a Text-to-Video model for coherent multi-shot 1080p videos with smooth, stable motion and faithful prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 lite Text to Video 480p

Seedance V1 Lite is a 480p Text-to-Video model for coherent multi-shot videos with smooth, stable motion and precise prompt following. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 lite Text to Video 720p

ByteDance Seedance V1 Lite produces coherent multi-shot 720p videos with smooth motion and accurate following of detailed text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro fast · image to video

Seedance V1 Pro Fast produces coherent multi-shot image-to-video with smooth, stable motion and faithful prompt control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro fast · text to video

Seedance V1 Pro Fast creates coherent multi-shot Text-to-Video outputs with smooth motion and strong prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro Image to Video 1080p

ByteDance Seedance V1 Pro generates coherent multi-shot image-to-video outputs with smooth motion and prompt-accurate results in 1080p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro Image to Video 480p

ByteDance Seedance v1 Pro i2v (480p) creates coherent multi-shot image-to-video sequences with smooth, stable motion and strong adherence to detailed prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro Image to Video 720p

ByteDance Seedance V1 Pro creates coherent multi-shot image-to-video clips with smooth, stable motion and faithful prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro Text to Video 1080p

ByteDance Seedance V1 Pro generates coherent multi-shot 1080p videos from text with smooth, stable motion and strong prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro Text to Video 480p

Seedance v1 Pro is a text-to-video 480p model for coherent multi-shot outputs with stable motion and accurate prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance v1 Pro Text to Video 720p

Seedance V1 Pro generates coherent multi-shot text-to-video at 720p with smooth, stable motion and precise prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

video upscaler

ByteDance Video Upscaler uses AI super-resolution to upscale videos to 4K and recover fine detail in a secure cloud environment. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

waver 1.0

Waver 1.0 by Bytedance is an all-in-one model for text-to-video, image-to-video, and text-to-image generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
character-ai

ovi · image to video

Ovi is a Veo-3-like image-to-video model that generates synchronized video and audio from text or text+image prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
character-ai

ovi · text to video

Ovi is a veo-3-like model that converts text or text+image prompts into synchronized video with audio. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
clarity-ai

crystal video upscaler

Clarity AI Crystal Video Upscaler increases video resolution with target megapixel control and Clarity AI crystal-video processing.

Video
Per video · quoted before submission
decart

lucy edit Pro

Lucy Edit Pro is a state-of-the-art video editing model that produces studio-quality results in minutes, not weeks. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
decart

lucy restyle

Lucy-Restyle is a state-of-the-art text-guided video editing model that transforms videos while preserving original motion, camera angles, and temporal consistency. Edit videos up to 10 minutes with natural language prompts. Ready-to-use REST inference API, ultra-fast processing, studio-grade quality.

Video
Per video · quoted before submission
elevenlabs

dubbing

ElevenLabs Dubbing automatically translates and dubs video/audio content into different languages while preserving the original speakers' voices. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
google

gemini omni 1.1 Flash · image to video

Gemini Omni 1.1 Flash Image-to-Video animates an input image into a short video with audio.

Video
Per video · quoted before submission
google

gemini omni 1.1 Flash · reference to video

Gemini Omni 1.1 Flash Reference-to-Video creates synchronized audiovisual clips from a prompt and optional image or video references.

Video
Per video · quoted before submission
google

gemini omni 1.1 Flash · text to video

Gemini Omni 1.1 Flash Text-to-Video creates short videos with synchronized audio from a text prompt.

Video
Per video · quoted before submission
google

gemini omni 1.1 Flash · video edit

Gemini Omni 1.1 Flash Video Edit applies natural-language edits to an existing video.

Video
Per video · quoted before submission
google

gemini omni Flash · image to video

Gemini Omni Flash Image-to-Video animates an input image into a short video with audio.

Video
Per video · quoted before submission
google

gemini omni Flash · reference to video

Gemini Omni Flash Reference-to-Video creates short videos from one or more reference images and a prompt.

Video
Per video · quoted before submission
google

gemini omni Flash · text to video

Gemini Omni Flash Text-to-Video creates short videos with synchronized audio from a text prompt.

Video
Per video · quoted before submission
google

gemini omni Flash · video edit

Gemini Omni Flash Video Edit applies natural-language edits to an existing video.

Video
Per video · quoted before submission
google

veo3.1 fast · image to video

Google Veo 3.1 Fast is an Image-to-Video model with native 1080p output for high-detail videos from images and fast performance. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3.1 fast · reference to video

Google Veo 3.1 Fast Reference-to-Video generates 8-second videos from up to three reference images using the official Veo predictLongRunning endpoint with referenceImages assets.

Video
Per video · quoted before submission
google

veo3.1 fast · text to video

Google Veo 3.1 Fast creates text-to-video with native 1080p and synchronized audio, delivering high-quality videos for creators. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3.1 fast · video extend

Extend Veo 3.1 videos in 7-second steps with the Fast endpoint—quick, coherent continuation that preserves style and motion, output as a single merged clip. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3.1 · image to video

Google Veo 3.1 is an Image-to-Video model that converts images into high-quality videos with native 1080P output for enhanced detail and creative flexibility. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3.1 lite · image to video

Google Veo 3.1 Lite transforms static images into high-fidelity videos with native audio.

Video
Per video · quoted before submission
google

veo3.1 lite · start end to video

Google Veo 3.1 Lite Start-End-to-Video generates videos by interpolating between start and end keyframe images.

Video
Per video · quoted before submission
google

veo3.1 lite · text to video

Google Veo 3.1 Lite generates high-fidelity videos with native audio from text prompts, optimized for cost efficiency.

Video
Per video · quoted before submission
google

veo3.1 · reference to video

Google Veo3.1 Reference-to-Video performs image-to-video generation that preserves a specific subject's appearance and identity from provided reference images, enabling consistent character or product motion across frames. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3.1 · text to video

Google Veo 3.1 converts text prompts into videos with synchronized audio at native 1080p for high-quality outputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3.1 · video extend

Extend and continue Veo 3.1 videos with smooth motion, preserved style, and strong scene coherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3 fast · image to video

Google Veo3 Fast provides faster, more cost-effective Image-to-Video generation vs Veo 3, with commercial use allowed and $0.25/sec pricing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
google

veo3 fast

Google Veo 3 Fast creates text-to-video with synchronized audio, delivering faster, more cost-effective results than standard Veo 3; commercial use allowed and pricing starts at $0.25/second. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission

Choose a model. Keep one video endpoint.

Pin a variant or let Auto compare compatible tasks within your key’s model pool and reservation limit.

Video API tutorial

Catalog prices refresh hourly. This public view may lag the latest sync by up to five minutes. Capabilities are provider-declared; availability can change.