THE MODEL CATALOG
Video model APIs & pricing
Find text-to-video, image-to-video and video processing APIs by task. Compare Wan versions, inspect each model’s parameters and submit asynchronous jobs with one Jevrouter key.
532 models · Page 10 of 12
v2.2 Text to Video
Pika v2.2 is a text-to-video model that creates high-quality videos from text prompts, supporting multiple video sizes and advanced prompt optimization. Ready-to-use REST API, no coldstarts, affordable pricing.
lipsync
PixVerse LipSync converts audio into realistic lip-sync animations with advanced algorithms for precise mouth movements and timing for video avatars. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
motion control · mimic
PixVerse Motion Control (Mimic) transfers the motion from a reference video onto a target image, creating a new video that follows the reference performance while preserving the target subject.
music mv agent
PixVerse Music MV Agent generates a music video from an input audio track with optional visual references, style controls, aspect ratio, and quality settings.
pixverse c1 · image to video
PixVerse C1 generates film-grade videos from a starting image with flexible duration (1-15s), multiple resolutions up to 1080p, and optional native audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse c1 · reference to video
PixVerse C1 Reference-to-Video generates videos from reference images with subject and background consistency.
pixverse c1 · text to video
PixVerse C1 generates film-grade videos from text prompts with flexible duration (1-15s), multiple resolutions up to 1080p, and optional native audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse c1 · transition
PixVerse C1 generates smooth transition videos between two images with flexible duration (1-15s), multiple resolutions up to 1080p, and optional native audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v4.5 Image to Video fast
Pixverse V4.5 I2V Fast converts images or text into high-quality videos with multi-resolution, aspect-ratio and motion-mode control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v4.5 Image to Video
PixVerse V4.5 I2V creates high-quality videos from text or image prompts, offering multiple resolutions, aspect ratios, and motion modes for versatile cinematic output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v4.5 Text to Video fast
Pixverse v4.5 Fast turns text prompts into high-quality videos with multiple resolutions, aspect ratios, and motion modes. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v4.5 Text to Video
Pixverse v4.5 turns text prompts into high-quality videos with multiple resolutions, aspect ratios, and motion modes. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v5.5 effects
PixVerse V5.5 Effects is an AI image-to-video model that converts still images into smooth, natural short videos with lifelike motion, supporting 5s/8s/10s clips at 360p to 1080p for social posts, ads, and previews. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v5.5 · image to video
PixVerse V5.5 Image-to-Video turns a single image into cinematic clips with smooth motion, clean detail, and strong subject fidelity—ideal for logo stingers, character motion, and social posts. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v5.5 · text to video
PixVerse V5.5 transforms text prompts into realistic videos with smooth motion and natural detail in seconds—ideal for stories, ads, and social clips. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v5.5 transition
Create smooth morph transitions between two images into 5s , 8s or 10s videos at 360p, 540p, 720p, or 1080p—perfect for logo reveals, before-and-after shots, and social posts. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v5.6 · image to video
PixVerse V5.5 Image-to-Video turns a single image into cinematic clips with smooth motion, clean detail, and strong subject fidelity—ideal for logo stingers, character motion, and social posts. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v5.6 · text to video
PixVerse V5.5 transforms text prompts into realistic videos with smooth motion and natural detail in seconds—ideal for stories, ads, and social clips. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v5 effects
PixVerse V5 Effects converts images into smooth, natural short videos with lifelike motion; supports 5s/8s and 720p/1080p outputs. Ready-to-use REST API, no coldstarts, best performance, affordable pricing.
pixverse v5 Image to Video
PixVerse V5 converts images to short, smooth, natural-looking videos. 5s video: $0.15 (360p/540p), $0.20 (720p), $0.40 (1080p). Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v5 Text to Video
PixVerse V5 Text-to-Video generates smooth, natural 5s videos from text prompts in seconds, with 720p output available ($0.20 per 5s). Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v5 transition
Create smooth morph transitions between two static images into 5s or 8s videos at 360p, 540p, 720p, or 1080p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
pixverse v6 · extend
PixVerse V6 Extend continues and enhances existing video content by extending the story forward using AI generation.
pixverse v6 · image to video
PixVerse V6 generates high-quality videos from images with flexible duration (1-15s), multiple resolutions up to 1080p, and optional audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v6 · reference to video
PixVerse V6 generates high-quality videos from images with flexible duration (1-15s), multiple resolutions up to 1080p, and optional audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v6 · text to video
PixVerse V6 generates high-quality videos from text prompts with flexible duration (1-15s), multiple resolutions up to 1080p, and optional audio generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
pixverse v6 · transition
PixVerse V6 Transition creates smooth AI-generated video transitions between a start image and an optional end image.
swap
PixVerse Swap replaces backgrounds, people, and objects directly inside existing videos for quick scene changes and creative edits with natural-looking results. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
p video 2 · image to video
Pruna P-Video-2 image-to-video generation with explicit duration, resolution, draft mode, and audio output.
p video 2 Pro · image to video
Pruna P-Video-2-Pro image-to-video with optional last-frame guidance generation at 480p or 768p, with generated audio.
p video 2 Pro · text to video
Pruna P-Video-2-Pro text-to-video generation at 480p or 768p, with generated audio.
p video 2 · text to video
Pruna P-Video-2 text-to-video generation with explicit duration, resolution, draft mode, and audio output.
p video · animate
Pruna p-video-animate model running on RunPod.
p video · avatar
Pruna p-video-avatar model running on RunPod.
p video · edit
Pruna P-Video-Edit edits an existing video from natural-language instructions, with optional reference images, prompt enhancement, draft generation, and source-audio preservation.
p video · image to video
Pruna p-video image-to-video model running on RunPod.
p video · replace
Pruna P-Video Replace replaces or inserts reference subjects into a source video while preserving motion, timing, and optional audio.
p video · text to video
Pruna p-video model running on RunPod.
aleph 2
Runway Aleph 2 is an in-context video editing model for precise prompt-based edits, multi-shot consistency, and optional keyframe guidance. Supports 2–30 second input videos and up to 5 keyframes at $1.85 per 5 seconds.
gen4 turbo
RunwayML Gen4 Turbo is an image-to-video model that generates high-quality videos from images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
upscale v1
RunwayML Upscale V1 upscales videos to 4K via simple file upload, billed at $0.11 per 5 seconds for fast, high-quality upscaling. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
skyreels v4 · image to video
SkyReels V4 Image to Video generates videos from image references and text prompts using the SkyReels V4 image2video workflow.
skyreels v4 · reference to video
SkyReels V4 Reference to Video generates videos from reference images or a reference video and a text prompt using the SkyReels V4 omni-video workflow.
skyreels v4 · text to video
SkyReels V4 Text to Video generates videos from text prompts using the SkyReels V4 text2video workflow.
lipsync 1.9.0 beta
Generate realistic lip-sync animations from audio using advanced algorithms for high-quality facial synchronization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
lipsync 2 Pro
Lipsync-2-pro creates studio-grade lip synchronization for video-to-video editing in minutes, not weeks. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
lipsync 2
Sync Lipsync-2 synchronizes lip movements in any video to supplied audio, enabling realistic mouth alignment for films, podcasts, games, or animations. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
lipsync 3 · avatar
Sync Lipsync 3 Avatar turns a still image into a talking character video driven by an input audio track.
Catalog prices refresh hourly. This public view may lag the latest sync by up to five minutes. Capabilities are provider-declared; availability can change.