Vidu Q1 Reference-to-video
Vidu Q1 Reference-to-Video is an advanced AI video generation model that brings static images to life. Upload a reference image and describe the motion you want — the model generates high-quality video with smooth animation, optional audio, and cinematic quality up to 1080p.

What it does
Vidu Q1 Reference-to-video, in practice.
Vidu Q1 Reference-to-Video is an efficient AI video generation model that generates video featuring specific subjects. Provide subject images alongside a motion prompt, and the model generates a 5-second 1080p video that faithfully preserves each subject's appearance, style, and identity — fast and at an accessible price point.
- Fast generation Optimized for quick turnaround with minimal wait time.
- Subject-driven generation Feature specific characters or objects with consistent appearance across the generated video.
- 1080p output Generate videos in full 1080p high definition quality.
- 5-second videos Produces crisp, fixed-length 5-second videos ready to share.
- Audio generation Optional audio with configurable type: full audio, speech only, or sound effects only.
- Prompt Enhancer Built-in tool to automatically improve your motion descriptions.
Run Vidu Q1 Reference-to-video
from $0.51/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
subjects | Information about the subjects in the images. Supports 1–7 subjects, total 1–7 images | |
prompt | Text prompt: A textual description for video generation, with a maximum length of 1500 characters. | |
duration | The duration of the generated media in seconds. Fixed at 5 for this model. | default 5 |
resolution | The resolution of the generated media. | 1080p |
aspect_ratio | The aspect ratio of the generated media. | 16:9 9:16 1:1 |
movement_amplitude | The movement amplitude of objects in the frame. | auto small medium large |
generate_audio | Whether to generate audio for the video. | default True |
audio_type | Audio type, required when audio is true, defaults to all. | all speech_only sound_effect_only |
seed | The random seed to use for the generation. -1 means a random seed will be used. | default |
FAQ
Short answers.
How much does Vidu Q1 Reference-to-video cost?
Pricing starts at $0.51 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $4.08. Usage is billed per request from your balance — no subscription.
Does Vidu Q1 Reference-to-video run uncensored here?
No. Vidu applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Vidu Q1 Reference-to-video for everything else it does well.
What does Vidu Q1 Reference-to-video take as input?
It is a text-to-video model. Vidu Q1 Reference-to-Video is an advanced AI video generation model that brings static images to life. Upload a reference image and describe the motion you want — the model generates high-quality video with smooth animation, optional audio, and cinematic quali
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related