THE MODEL CATALOG
Video model APIs & pricing
Find text-to-video, image-to-video and video processing APIs by task. Compare Wan versions, inspect each model’s parameters and submit asynchronous jobs with one Jevrouter key.
532 models · Page 2 of 12
flux 3 · video edit
FLUX 3 Video Edit applies prompt-guided changes to an existing video while preserving its motion, timing, and framing.
flux 3 · video extend draft
Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .
flux 3 · video extend
Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .
flux 3 · video upscale
FLUX Video Upscale enhances an input video with source-faithful or creative detail reconstruction while preserving its aspect ratio.
fibo · video background remover
Bria Video Background Remover removes the background from videos with support for transparency and custom background colors. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
fibo · video upscaler
Bria Video Upscaler increases video resolution up to 8K with 2x or 4x upscaling. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
video eraser · mask
AI-powered video object eraser that removes unwanted objects from videos using mask videos.
video eraser · prompt
AI-powered video object eraser that removes unwanted objects from videos based on text prompts.
avatar omni human 1.5
OmniHuman 1.5 converts audio and visual cues into lifelike avatar animations for virtual humans, storytelling, and interactive agents. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
avatar omni human
Bytedance OmniHuman turns a single portrait photo into avatar video with lifelike motion and expressions ($0.12/sec). Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
dreamactor v2
DreamActor V2 transfers motion from a driving video to characters in an image. Great performance for non-human and multiple characters.
latentsync
Bytedance LatentSync combines Stable Diffusion and TREPA for high-res end-to-end lip-sync, delivering precise, realistic mouth motions in generated videos. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
lipsync · audio to video
Bytedance LipSync turns audio into lifelike talking videos by generating precise lip movements fully synced to input audio. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
seedance 2.0 fast · image to video spicy
Seedance 2.0 Fast Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with faster turnaround and lower cost.
seedance 2.0 fast · image to video turbo
Seedance 2.0 Fast (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts using speed-optimized inference —the fastest Seedance image-to-video option with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture.
seedance 2.0 fast · image to video
Seedance 2.0 Fast (Image-to-Video) generates cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level control, and exceptional motion stability — optimized for faster generation. Built on ByteDance Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive motion.
seedance 2.0 fast · text to video turbo
Seedance 2.0 Fast (Text-to-Video Turbo) generates cinematic 720p/1080p videos from text prompts using speed-optimized inference —the fastest Seedance option with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
seedance 2.0 fast · text to video
Seedance 2.0 Fast (Text-to-Video) generates cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for faster generation. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
seedance 2.0 fast · video edit turbo
Seedance 2.0 Fast (Video-Edit Turbo) is the fastest, cheapest turbo tier for editing an input video from a natural-language prompt — high-resolution output with optimized cost and speed.
seedance 2.0 fast · video edit
Seedance 2.0 Fast (Video-Edit) edits an input video from a natural-language prompt at a faster, cheaper tier. Built on ByteDance Seed's unified multimodal architecture, it preserves subject identity, composition, and motion while rewriting lighting, style, weather, environment, or specific elements as instructed.
seedance 2.0 fast · video extend
Seedance 2.0 Fast (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt — at the faster, cheaper Seedance 2.0 Fast tier.
seedance 2.0 · image to video spicy
Seedance 2.0 Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with smooth animations and stable aesthetics.
seedance 2.0 · image to video turbo
Seedance 2.0 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture.
seedance 2.0 · image to video
Seedance 2.0 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.
seedance 2.0 mini · image to video spicy
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · image to video turbo
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · image to video
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · text to video turbo
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · text to video
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · video edit turbo
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · video edit
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 mini · video extend
Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16. Priced at 50% of standard Seedance 2.0.
seedance 2.0 · text to video turbo
Seedance 2.0 (Text-to-Video Turbo) generates cinematic videos from text prompts at 720p and 1080p with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for turbo output. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
seedance 2.0 · text to video
Seedance 2.0 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
seedance 2.0 · video edit turbo
Seedance 2.0 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion.
seedance 2.0 · video edit
Seedance 2.0 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output.
seedance 2.0 · video extend
Seedance 2.0 (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt.
seedance 2.5 · image to video spicy
Seedance 2.5 Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with smooth animations and stable aesthetics.
seedance 2.5 · image to video turbo
Seedance 2.5 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture.
seedance 2.5 · image to video
Seedance 2.5 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.
seedance 2.5 · talking avatar
Animate a portrait image from an audio recording with Seedance 2.5. Generate talking-avatar videos in 480p or 720p, processing up to the first 120 seconds of audio.
seedance 2.5 · text to video turbo
Seedance 2.5 (Text-to-Video Turbo) generates cinematic videos from text prompts at 720p and 1080p with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for turbo output. Built on ByteDance Seed's unified multimodal architecture. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
seedance 2.5 · text to video
Seedance 2.5 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
seedance 2.5 · video edit turbo
Seedance 2.5 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion.
seedance 2.5 · video edit
Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output.
seedance 2.5 · video extend
Seedance 2.5 (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt.
seedance v1.5 Pro · image to video fast
Seedance 1.5 Pro Fast (Image-to-Video) generates cinematic, live-action–leaning clips from a text prompt plus a first-frame image, preserving the image's subject and composition while adding expressive motion and stable aesthetics. It supports 4–12s duration control, adaptive aspect ratio that follows the input image, and reproducible outputs via seeds—ideal for ad creatives and short-drama shots that need a strong visual anchor.
seedance v1.5 Pro · image to video spicy
Seedance 1.5 Pro Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a text prompt plus a first-frame image, optimized for scalable content generation with smooth animations and stable aesthetics.
Catalog prices refresh hourly. This public view may lag the latest sync by up to five minutes. Capabilities are provider-declared; availability can change.