Vidu Q2-Pro Reference-to-video
Vidu Q2-Pro Reference-to-Video is an advanced AI video generation model that brings static images to life. Upload a reference image and describe the motion you want — the model generates high-quality video with smooth animation, optional audio, and cinematic quality up to 1080p.
What it does
Vidu Q2-Pro Reference-to-video, in practice.
Vidu Q2-Pro Reference-to-Video is a professional-grade AI video generation model that generates video featuring specific subjects with cinematic precision. Provide subject images alongside a motion prompt, and the model delivers up to 1080p video with rich detail, strict subject fidelity, and smooth natural motion — ideal for high-end creative, brand, and production workflows.
- Professional quality Cinematic detail and smooth motion with faithful subject preservation at up to 1080p.
- Subject-driven generation Feature specific characters or objects with strict visual fidelity throughout the video.
- Flexible duration Create videos up to 10 seconds in length.
- Audio generation Optional audio with configurable type: full audio, speech only, or sound effects only.
- Motion control Adjust movement amplitude for subtle or dynamic animations.
- Prompt Enhancer Built-in tool to automatically improve your motion descriptions.
Run Vidu Q2-Pro Reference-to-video
from $0.13/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
subjects | Information about the subjects in the images. Supports 1–7 subjects, total 1–7 images | |
prompt | Text prompt: A textual description for video generation, with a maximum length of 1500 characters. | |
generate_audio | Whether to generate audio. | default True |
bgm | The background music for generating the output. | default True |
audio_type | Audio type, required when audio is true, defaults to all. | all speech_only sound_effect_only |
duration | The duration of the generated media in seconds. | default 5 |
seed | The random seed to use for the generation. | default |
aspect_ratio | The aspect ratio of the output video. Defaults to 16:9, accepted: 16:9 9:16 1:1. | 16:9 9:16 1:1 |
resolution | The resolution of the generated media. | 540p 720p 1080p |
Sample prompt
The prompt behind the sample.
Remove the figures from the picture.
generate_audio: Truebgm: Trueaudio_type: allduration: 5seed: aspect_ratio: 16:9resolution: 720pFAQ
Short answers.
How much does Vidu Q2-Pro Reference-to-video cost?
Pricing starts at $0.13 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $1.02. Usage is billed per request from your balance — no subscription.
Does Vidu Q2-Pro Reference-to-video run uncensored here?
No. Vidu applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Vidu Q2-Pro Reference-to-video for everything else it does well.
What does Vidu Q2-Pro Reference-to-video take as input?
It is a text-to-video model. Vidu Q2-Pro Reference-to-Video is an advanced AI video generation model that brings static images to life. Upload a reference image and describe the motion you want — the model generates high-quality video with smooth animation, optional audio, and cinematic q
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related