minimaxiH3 Get access

Models

Video-to-video models.

Extend, edit, restyle or re-time an existing clip. Models that take video in and give video back. 18 models, priced per second, with a sample output each.

All 328Text-to-video 71Image-to-video 111Video-to-video 18Audio-to-video 3Text-to-image 60Image editing 65
Google · Video-to-video

Gemini Omni 1.1 Flash Video Extend

A natively multimodal Google DeepMind model that continues an existing clip with a seamlessly matched 3-to-10-second ext…

from $0.055/sec
Google · Video-to-video

Gemini Omni 1.1 Flash Video Edit

A natively multimodal Google DeepMind model that applies a text-instructed edit to an existing video - adding, removing,…

from $0.055/sec
Qwen · Video-to-video

Wan-3.0-Prime Reference-to-video

All-in-One Reference: keep subjects consistent from any mix of reference images, videos, and audio; pixel-level identity…

from $0.091/sec
Qwen · Video-to-video

Wan-3.0 Reference-to-video

All-in-One Reference: keep subjects consistent from any mix of reference images, videos, and audio; pixel-level identity…

from $0.060/sec
Google · Video-to-video

Gemini Omni Flash Video Edit

A natively multimodal Google DeepMind model that edits an existing video from a text prompt with optional reference imag…

from $0.21/sec
Google · Video-to-video

Gemini Omni Flash Reference-to-Video Developer

Gemini Omni Flash is Google's multimodal video generation model. This reference-to-video variant transforms existing vid…

from $0.18/sec
Qwen · Video-to-video

HappyHorse-1.0 Video-edit

Edits an input video with text instructions and optional reference images, supporting 720P or 1080P output.

from $0.21/sec
Qwen · Video-to-video

Wan-2.7 Reference-to-video

Generates character-driven videos from reference images and videos, with multi-subject and voice-cloning support.

from $0.15/sec
Qwen · Video-to-video

Wan-2.7 Video-edit

Edits videos using text instructions, reference images, and style transfer with multi-modal input support.

from $0.15/sec
Qwen · Video-to-video

Wan-2.6 Video-to-video

A speed-optimized video-to-video option that prioritizes lower latency while retaining strong visual fidelity. Ideal for…

from $0.11/sec
Kuaishou · Video-to-video

Kling Video O3 Pro Video-Edit

Kling Omni Video O3 Video-Edit enables conversational video editing through natural language commands. Professional qual…

from $0.21/sec
Kuaishou · Video-to-video

Kling Video O3 Std Video-Edit

Kling Omni Video O3 Video-Edit (Standard) enables natural-language video edits: remove or replace objects, change backgr…

from $0.16/sec
PixVerse · Video-to-video

Pixverse v6 Video-Extend

Pixverse v6 Video Extend model. High-quality video generation from image prompts.

from $0.038/sec
BytePlus · Video-to-video

BytePlus Video Upscaler

BytePlus AI MediaKit video quality enhancement (super-resolution): upscale/enhance a source video to up to 8K with scene…

from $0.0045/sec
Lightricks · Video-to-video

Ltx 2.3 Quality Extend Video

Generate high-quality video with audio from images using LTX-2.3

from $0.0030/sec
Sync · Video-to-video

Sync.so Lipsync v3

Sync.so Lipsync v3 (sync-3) is Sync Labs state-of-the-art lip-sync model, re-syncing the lips of an existing video to a …

from $0.33/sec
VEED · Video-to-video

VEED Lipsync

VEED Lipsync re-drives the lip movements of an existing talking-head video to match a new audio track, preserving identi…

from $0.019/sec
Tencent · Video-to-video

Tencent Video Upscaler

Tencent Video Upscaler (MPS super-resolution / enhancement)

from $0.0060/sec

Run any of them in Chat.

Invite code opens Chat with every model above. No code? Join the waitlist and tell us which model you need.

Get access Browse models