THE MODEL CATALOG

Video model APIs & pricing

Find text-to-video, image-to-video and video processing APIs by task. Compare Wan versions, inspect each model’s parameters and submit asynchronous jobs with one Jevrouter key.

532 models · Page 10 of 12

pika

v2.2 Text to Video

Pika v2.2 is a text-to-video model that creates high-quality videos from text prompts, supporting multiple video sizes and advanced prompt optimization. Ready-to-use REST API, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

lipsync

PixVerse LipSync converts audio into realistic lip-sync animations with advanced algorithms for precise mouth movements and timing for video avatars. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

motion control · mimic

PixVerse Motion Control (Mimic) transfers the motion from a reference video onto a target image, creating a new video that follows the reference performance while preserving the target subject.

Video
Per video · quoted before submission
pixverse

music mv agent

PixVerse Music MV Agent generates a music video from an input audio track with optional visual references, style controls, aspect ratio, and quality settings.

Video
Per video · quoted before submission
pixverse

pixverse c1 · image to video

PixVerse C1 generates film-grade videos from a starting image with flexible duration (1-15s), multiple resolutions up to 1080p, and optional native audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse c1 · reference to video

PixVerse C1 Reference-to-Video generates videos from reference images with subject and background consistency.

Video
Per video · quoted before submission
pixverse

pixverse c1 · text to video

PixVerse C1 generates film-grade videos from text prompts with flexible duration (1-15s), multiple resolutions up to 1080p, and optional native audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse c1 · transition

PixVerse C1 generates smooth transition videos between two images with flexible duration (1-15s), multiple resolutions up to 1080p, and optional native audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v4.5 Image to Video fast

Pixverse V4.5 I2V Fast converts images or text into high-quality videos with multi-resolution, aspect-ratio and motion-mode control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v4.5 Image to Video

PixVerse V4.5 I2V creates high-quality videos from text or image prompts, offering multiple resolutions, aspect ratios, and motion modes for versatile cinematic output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v4.5 Text to Video fast

Pixverse v4.5 Fast turns text prompts into high-quality videos with multiple resolutions, aspect ratios, and motion modes. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v4.5 Text to Video

Pixverse v4.5 turns text prompts into high-quality videos with multiple resolutions, aspect ratios, and motion modes. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5.5 effects

PixVerse V5.5 Effects is an AI image-to-video model that converts still images into smooth, natural short videos with lifelike motion, supporting 5s/8s/10s clips at 360p to 1080p for social posts, ads, and previews. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5.5 · image to video

PixVerse V5.5 Image-to-Video turns a single image into cinematic clips with smooth motion, clean detail, and strong subject fidelity—ideal for logo stingers, character motion, and social posts. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5.5 · text to video

PixVerse V5.5 transforms text prompts into realistic videos with smooth motion and natural detail in seconds—ideal for stories, ads, and social clips. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5.5 transition

Create smooth morph transitions between two images into 5s , 8s or 10s videos at 360p, 540p, 720p, or 1080p—perfect for logo reveals, before-and-after shots, and social posts. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5.6 · image to video

PixVerse V5.5 Image-to-Video turns a single image into cinematic clips with smooth motion, clean detail, and strong subject fidelity—ideal for logo stingers, character motion, and social posts. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5.6 · text to video

PixVerse V5.5 transforms text prompts into realistic videos with smooth motion and natural detail in seconds—ideal for stories, ads, and social clips. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5 effects

PixVerse V5 Effects converts images into smooth, natural short videos with lifelike motion; supports 5s/8s and 720p/1080p outputs. Ready-to-use REST API, no coldstarts, best performance, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5 Image to Video

PixVerse V5 converts images to short, smooth, natural-looking videos. 5s video: $0.15 (360p/540p), $0.20 (720p), $0.40 (1080p). Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5 Text to Video

PixVerse V5 Text-to-Video generates smooth, natural 5s videos from text prompts in seconds, with 720p output available ($0.20 per 5s). Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v5 transition

Create smooth morph transitions between two static images into 5s or 8s videos at 360p, 540p, 720p, or 1080p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v6 · extend

PixVerse V6 Extend continues and enhances existing video content by extending the story forward using AI generation.

Video
Per video · quoted before submission
pixverse

pixverse v6 · image to video

PixVerse V6 generates high-quality videos from images with flexible duration (1-15s), multiple resolutions up to 1080p, and optional audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v6 · reference to video

PixVerse V6 generates high-quality videos from images with flexible duration (1-15s), multiple resolutions up to 1080p, and optional audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v6 · text to video

PixVerse V6 generates high-quality videos from text prompts with flexible duration (1-15s), multiple resolutions up to 1080p, and optional audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pixverse

pixverse v6 · transition

PixVerse V6 Transition creates smooth AI-generated video transitions between a start image and an optional end image.

Video
Per video · quoted before submission
pixverse

swap

PixVerse Swap replaces backgrounds, people, and objects directly inside existing videos for quick scene changes and creative edits with natural-looking results. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Video
Per video · quoted before submission
pruna-ai

p video 2 · image to video

Pruna P-Video-2 image-to-video generation with explicit duration, resolution, draft mode, and audio output.

Video
Per video · quoted before submission
pruna-ai

p video 2 Pro · image to video

Pruna P-Video-2-Pro image-to-video with optional last-frame guidance generation at 480p or 768p, with generated audio.

Video
Per video · quoted before submission
pruna-ai

p video 2 Pro · text to video

Pruna P-Video-2-Pro text-to-video generation at 480p or 768p, with generated audio.

Video
Per video · quoted before submission
pruna-ai

p video 2 · text to video

Pruna P-Video-2 text-to-video generation with explicit duration, resolution, draft mode, and audio output.

Video
Per video · quoted before submission
pruna-ai

p video · animate

Pruna p-video-animate model running on RunPod.

Video
Per video · quoted before submission
pruna-ai

p video · avatar

Pruna p-video-avatar model running on RunPod.

Video
Per video · quoted before submission
pruna-ai

p video · edit

Pruna P-Video-Edit edits an existing video from natural-language instructions, with optional reference images, prompt enhancement, draft generation, and source-audio preservation.

Video
Per video · quoted before submission
pruna-ai

p video · image to video

Pruna p-video image-to-video model running on RunPod.

Video
Per video · quoted before submission
pruna-ai

p video · replace

Pruna P-Video Replace replaces or inserts reference subjects into a source video while preserving motion, timing, and optional audio.

Video
Per video · quoted before submission
pruna-ai

p video · text to video

Pruna p-video model running on RunPod.

Video
Per video · quoted before submission
runwayml

aleph 2

Runway Aleph 2 is an in-context video editing model for precise prompt-based edits, multi-shot consistency, and optional keyframe guidance. Supports 2–30 second input videos and up to 5 keyframes at $1.85 per 5 seconds.

Video
Per video · quoted before submission
runwayml

gen4 turbo

RunwayML Gen4 Turbo is an image-to-video model that generates high-quality videos from images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
runwayml

upscale v1

RunwayML Upscale V1 upscales videos to 4K via simple file upload, billed at $0.11 per 5 seconds for fast, high-quality upscaling. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
skywork-ai

skyreels v4 · image to video

SkyReels V4 Image to Video generates videos from image references and text prompts using the SkyReels V4 image2video workflow.

Video
Per video · quoted before submission
skywork-ai

skyreels v4 · reference to video

SkyReels V4 Reference to Video generates videos from reference images or a reference video and a text prompt using the SkyReels V4 omni-video workflow.

Video
Per video · quoted before submission
skywork-ai

skyreels v4 · text to video

SkyReels V4 Text to Video generates videos from text prompts using the SkyReels V4 text2video workflow.

Video
Per video · quoted before submission
sync

lipsync 1.9.0 beta

Generate realistic lip-sync animations from audio using advanced algorithms for high-quality facial synchronization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
sync

lipsync 2 Pro

Lipsync-2-pro creates studio-grade lip synchronization for video-to-video editing in minutes, not weeks. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
sync

lipsync 2

Sync Lipsync-2 synchronizes lip movements in any video to supplied audio, enabling realistic mouth alignment for films, podcasts, games, or animations. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Video
Per video · quoted before submission
sync

lipsync 3 · avatar

Sync Lipsync 3 Avatar turns a still image into a talking character video driven by an input audio track.

Video
Per video · quoted before submission

Choose a model. Keep one video endpoint.

Pin a variant or let Auto compare compatible tasks within your key’s model pool and reservation limit.

Video API tutorial

Catalog prices refresh hourly. This public view may lag the latest sync by up to five minutes. Capabilities are provider-declared; availability can change.