THE MODEL CATALOG
Video model APIs & pricing
Find text-to-video, image-to-video and video processing APIs by task. Compare Wan versions, inspect each model’s parameters and submit asynchronous jobs with one Jevrouter key.
532 models · Page 3 of 12
seedance v1.5 Pro · image to video
Seedance 1.5 Pro (Image-to-Video) generates cinematic, live-action–leaning clips from a text prompt plus a first-frame image, preserving the image’s subject and composition while adding expressive motion and stable aesthetics. It supports 4–12s duration control (including Smart Duration), adaptive aspect ratio that follows the input image, and reproducible outputs via seeds—ideal for ad creatives and short-drama shots that need a strong visual anchor.
seedance v1.5 Pro · text to video fast
Seedance 1.5 Pro Fast (Text-to-Video) generates cinematic, live-action–leaning clips from text with strong prompt adherence, expressive motion, and stable aesthetics. It supports 4–12s duration control, multiple aspect ratios, and reproducible generation via seeds—ideal for ads and short-drama workflows.
seedance v1.5 Pro · text to video
Seedance 1.5 Pro (Text-to-Video) generates cinematic, live-action–leaning clips from text with strong prompt adherence, expressive motion, and stable aesthetics. It supports 4–12s duration control (including Smart Duration), multiple aspect ratios (including adaptive), and reproducible generation via seeds—ideal for ads and short-drama workflows.
seedance v1.5 Pro · video extend fast
Seedance 1.5 Pro Fast Video-Extend turns short video clips into longer videos with natural motion continuation, stable aesthetics, and upscaled output. It supports 4-12s duration control, 720p/1080p resolutions, and reproducible generation via seeds—ideal for extending ad creatives and short-drama shots with higher quality output.
seedance v1.5 Pro · video extend
Seedance 1.5 Pro Video-Extend turns short video clips into longer videos with natural motion continuation and stable aesthetics. It supports 4-12s duration control, multiple resolutions, and reproducible generation via seeds—ideal for extending ad creatives and short-drama shots.
seedance v1 lite Image to Video 1080p
Seedance Lite creates coherent multi-shot image-to-video clips in 1080p with smooth, stable motion and accurate prompt following. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 lite Image to Video 480p
ByteDance Seedance V1 Lite is a 480p Image-to-Video model for coherent multi-shot videos with smooth motion and precise prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 lite Image to Video 720p
ByteDance Seedance Lite i2v 720p creates coherent multi-shot image-to-video clips with smooth, stable motion and prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 lite Text to Video 1080p
Seedance v1 Lite is a Text-to-Video model for coherent multi-shot 1080p videos with smooth, stable motion and faithful prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 lite Text to Video 480p
Seedance V1 Lite is a 480p Text-to-Video model for coherent multi-shot videos with smooth, stable motion and precise prompt following. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 lite Text to Video 720p
ByteDance Seedance V1 Lite produces coherent multi-shot 720p videos with smooth motion and accurate following of detailed text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro fast · image to video
Seedance V1 Pro Fast produces coherent multi-shot image-to-video with smooth, stable motion and faithful prompt control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro fast · text to video
Seedance V1 Pro Fast creates coherent multi-shot Text-to-Video outputs with smooth motion and strong prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro Image to Video 1080p
ByteDance Seedance V1 Pro generates coherent multi-shot image-to-video outputs with smooth motion and prompt-accurate results in 1080p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro Image to Video 480p
ByteDance Seedance v1 Pro i2v (480p) creates coherent multi-shot image-to-video sequences with smooth, stable motion and strong adherence to detailed prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro Image to Video 720p
ByteDance Seedance V1 Pro creates coherent multi-shot image-to-video clips with smooth, stable motion and faithful prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro Text to Video 1080p
ByteDance Seedance V1 Pro generates coherent multi-shot 1080p videos from text with smooth, stable motion and strong prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro Text to Video 480p
Seedance v1 Pro is a text-to-video 480p model for coherent multi-shot outputs with stable motion and accurate prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance v1 Pro Text to Video 720p
Seedance V1 Pro generates coherent multi-shot text-to-video at 720p with smooth, stable motion and precise prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
video upscaler
ByteDance Video Upscaler uses AI super-resolution to upscale videos to 4K and recover fine detail in a secure cloud environment. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
waver 1.0
Waver 1.0 by Bytedance is an all-in-one model for text-to-video, image-to-video, and text-to-image generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
ovi · image to video
Ovi is a Veo-3-like image-to-video model that generates synchronized video and audio from text or text+image prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
ovi · text to video
Ovi is a veo-3-like model that converts text or text+image prompts into synchronized video with audio. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
crystal video upscaler
Clarity AI Crystal Video Upscaler increases video resolution with target megapixel control and Clarity AI crystal-video processing.
lucy edit Pro
Lucy Edit Pro is a state-of-the-art video editing model that produces studio-quality results in minutes, not weeks. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
lucy restyle
Lucy-Restyle is a state-of-the-art text-guided video editing model that transforms videos while preserving original motion, camera angles, and temporal consistency. Edit videos up to 10 minutes with natural language prompts. Ready-to-use REST inference API, ultra-fast processing, studio-grade quality.
dubbing
ElevenLabs Dubbing automatically translates and dubs video/audio content into different languages while preserving the original speakers' voices. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
gemini omni 1.1 Flash · image to video
Gemini Omni 1.1 Flash Image-to-Video animates an input image into a short video with audio.
gemini omni 1.1 Flash · reference to video
Gemini Omni 1.1 Flash Reference-to-Video creates synchronized audiovisual clips from a prompt and optional image or video references.
gemini omni 1.1 Flash · text to video
Gemini Omni 1.1 Flash Text-to-Video creates short videos with synchronized audio from a text prompt.
gemini omni 1.1 Flash · video edit
Gemini Omni 1.1 Flash Video Edit applies natural-language edits to an existing video.
gemini omni Flash · image to video
Gemini Omni Flash Image-to-Video animates an input image into a short video with audio.
gemini omni Flash · reference to video
Gemini Omni Flash Reference-to-Video creates short videos from one or more reference images and a prompt.
gemini omni Flash · text to video
Gemini Omni Flash Text-to-Video creates short videos with synchronized audio from a text prompt.
gemini omni Flash · video edit
Gemini Omni Flash Video Edit applies natural-language edits to an existing video.
veo3.1 fast · image to video
Google Veo 3.1 Fast is an Image-to-Video model with native 1080p output for high-detail videos from images and fast performance. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3.1 fast · reference to video
Google Veo 3.1 Fast Reference-to-Video generates 8-second videos from up to three reference images using the official Veo predictLongRunning endpoint with referenceImages assets.
veo3.1 fast · text to video
Google Veo 3.1 Fast creates text-to-video with native 1080p and synchronized audio, delivering high-quality videos for creators. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3.1 fast · video extend
Extend Veo 3.1 videos in 7-second steps with the Fast endpoint—quick, coherent continuation that preserves style and motion, output as a single merged clip. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3.1 · image to video
Google Veo 3.1 is an Image-to-Video model that converts images into high-quality videos with native 1080P output for enhanced detail and creative flexibility. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3.1 lite · image to video
Google Veo 3.1 Lite transforms static images into high-fidelity videos with native audio.
veo3.1 lite · start end to video
Google Veo 3.1 Lite Start-End-to-Video generates videos by interpolating between start and end keyframe images.
veo3.1 lite · text to video
Google Veo 3.1 Lite generates high-fidelity videos with native audio from text prompts, optimized for cost efficiency.
veo3.1 · reference to video
Google Veo3.1 Reference-to-Video performs image-to-video generation that preserves a specific subject's appearance and identity from provided reference images, enabling consistent character or product motion across frames. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3.1 · text to video
Google Veo 3.1 converts text prompts into videos with synchronized audio at native 1080p for high-quality outputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3.1 · video extend
Extend and continue Veo 3.1 videos with smooth motion, preserved style, and strong scene coherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3 fast · image to video
Google Veo3 Fast provides faster, more cost-effective Image-to-Video generation vs Veo 3, with commercial use allowed and $0.25/sec pricing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
veo3 fast
Google Veo 3 Fast creates text-to-video with synchronized audio, delivering faster, more cost-effective results than standard Veo 3; commercial use allowed and pricing starts at $0.25/second. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Catalog prices refresh hourly. This public view may lag the latest sync by up to five minutes. Capabilities are provider-declared; availability can change.