THE MODEL CATALOG

Video model APIs & pricing

Find text-to-video, image-to-video and video processing APIs by task. Compare Wan versions, inspect each model’s parameters and submit asynchronous jobs with one Jevrouter key.

532 models · Page 2 of 12

black-forest-labs

flux 3 · video edit

FLUX 3 Video Edit applies prompt-guided changes to an existing video while preserving its motion, timing, and framing.

Video
Per video · quoted before submission
black-forest-labs

flux 3 · video extend draft

Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .

Video
Per video · quoted before submission
black-forest-labs

flux 3 · video extend

Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .

Video
Per video · quoted before submission
black-forest-labs

flux 3 · video upscale

FLUX Video Upscale enhances an input video with source-faithful or creative detail reconstruction while preserving its aspect ratio.

Video
Per video · quoted before submission
bria

fibo · video background remover

Bria Video Background Remover removes the background from videos with support for transparency and custom background colors. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bria

fibo · video upscaler

Bria Video Upscaler increases video resolution up to 8K with 2x or 4x upscaling. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bria

video eraser · mask

AI-powered video object eraser that removes unwanted objects from videos using mask videos.

Video
Per video · quoted before submission
bria

video eraser · prompt

AI-powered video object eraser that removes unwanted objects from videos based on text prompts.

Video
Per video · quoted before submission
bytedance

avatar omni human 1.5

OmniHuman 1.5 converts audio and visual cues into lifelike avatar animations for virtual humans, storytelling, and interactive agents. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

avatar omni human

Bytedance OmniHuman turns a single portrait photo into avatar video with lifelike motion and expressions ($0.12/sec). Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

dreamactor v2

DreamActor V2 transfers motion from a driving video to characters in an image. Great performance for non-human and multiple characters.

Video
Per video · quoted before submission
bytedance

latentsync

Bytedance LatentSync combines Stable Diffusion and TREPA for high-res end-to-end lip-sync, delivering precise, realistic mouth motions in generated videos. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

lipsync · audio to video

Bytedance LipSync turns audio into lifelike talking videos by generating precise lip movements fully synced to input audio. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · image to video spicy

Seedance 2.0 Fast Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with faster turnaround and lower cost.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · image to video turbo

Seedance 2.0 Fast (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts using speed-optimized inference —the fastest Seedance image-to-video option with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · image to video

Seedance 2.0 Fast (Image-to-Video) generates cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level control, and exceptional motion stability — optimized for faster generation. Built on ByteDance Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive motion.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · text to video turbo

Seedance 2.0 Fast (Text-to-Video Turbo) generates cinematic 720p/1080p videos from text prompts using speed-optimized inference —the fastest Seedance option with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · text to video

Seedance 2.0 Fast (Text-to-Video) generates cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for faster generation. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · video edit turbo

Seedance 2.0 Fast (Video-Edit Turbo) is the fastest, cheapest turbo tier for editing an input video from a natural-language prompt — high-resolution output with optimized cost and speed.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · video edit

Seedance 2.0 Fast (Video-Edit) edits an input video from a natural-language prompt at a faster, cheaper tier. Built on ByteDance Seed's unified multimodal architecture, it preserves subject identity, composition, and motion while rewriting lighting, style, weather, environment, or specific elements as instructed.

Video
Per video · quoted before submission
bytedance

seedance 2.0 fast · video extend

Seedance 2.0 Fast (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt — at the faster, cheaper Seedance 2.0 Fast tier.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · image to video spicy

Seedance 2.0 Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with smooth animations and stable aesthetics.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · image to video turbo

Seedance 2.0 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · image to video

Seedance 2.0 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · image to video spicy

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · image to video turbo

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · image to video

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · text to video turbo

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · text to video

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · video edit turbo

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · video edit

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 mini · video extend

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · text to video turbo

Seedance 2.0 (Text-to-Video Turbo) generates cinematic videos from text prompts at 720p and 1080p with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for turbo output. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · text to video

Seedance 2.0 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · video edit turbo

Seedance 2.0 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · video edit

Seedance 2.0 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output.

Video
Per video · quoted before submission
bytedance

seedance 2.0 · video extend

Seedance 2.0 (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · image to video spicy

Seedance 2.5 Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with smooth animations and stable aesthetics.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · image to video turbo

Seedance 2.5 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · image to video

Seedance 2.5 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · talking avatar

Animate a portrait image from an audio recording with Seedance 2.5. Generate talking-avatar videos in 480p or 720p, processing up to the first 120 seconds of audio.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · text to video turbo

Seedance 2.5 (Text-to-Video Turbo) generates cinematic videos from text prompts at 720p and 1080p with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for turbo output. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · text to video

Seedance 2.5 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · video edit turbo

Seedance 2.5 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · video edit

Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output.

Video
Per video · quoted before submission
bytedance

seedance 2.5 · video extend

Seedance 2.5 (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt.

Video
Per video · quoted before submission
bytedance

seedance v1.5 Pro · image to video fast

Seedance 1.5 Pro Fast (Image-to-Video) generates cinematic, live-action–leaning clips from a text prompt plus a first-frame image, preserving the image's subject and composition while adding expressive motion and stable aesthetics. It supports 4–12s duration control, adaptive aspect ratio that follows the input image, and reproducible outputs via seeds—ideal for ad creatives and short-drama shots that need a strong visual anchor.

Video
Per video · quoted before submission
bytedance

seedance v1.5 Pro · image to video spicy

Seedance 1.5 Pro Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a text prompt plus a first-frame image, optimized for scalable content generation with smooth animations and stable aesthetics.

Video
Per video · quoted before submission

Choose a model. Keep one video endpoint.

Pin a variant or let Auto compare compatible tasks within your key’s model pool and reservation limit.

Video API tutorial

Catalog prices refresh hourly. This public view may lag the latest sync by up to five minutes. Capabilities are provider-declared; availability can change.