Vidu Q2 Text-to-video
Vidu Q2 Text-to-Video is an advanced AI video generation model that brings static images to life. Upload a reference image and describe the motion you want — the model generates high-quality video with smooth animation, optional audio, and cinematic quality up to 1080p.
What it does
Vidu Q2 Text-to-video, in practice.
Vidu Q2 Text-to-Video is a capable AI video generation model that creates videos directly from text descriptions. Positioned between the entry-level Q1 and the professional Q2-Pro, it delivers a strong balance of quality, speed, and affordability — suitable for a wide range of creative and commercial workflows.
- Balanced quality and speed Strong visual output without the wait time of higher-tier models.
- High resolution output Generate videos in 540p, 720p, or 1080p quality.
- Flexible duration Create videos from 1 to 10 seconds in length.
- Audio generation Optional synchronized audio and background music.
- Motion control Adjust movement amplitude for subtle or dynamic animations.
- Prompt Enhancer Built-in tool to automatically improve your video descriptions.
Run Vidu Q2 Text-to-video
from $0.063/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | The positive prompt for the generation. | |
style | The style of output video. | general anime |
resolution | The resolution of the generated media. | 540p 720p 1080p |
duration | The duration of the generated media in seconds. | default 5 |
generate_audio | Whether to generate audio for the video. | default True |
aspect_ratio | The aspect ratio of the generated media. | 16:9 9:16 1:1 4:3 3:4 |
movement_amplitude | The movement amplitude of objects in the frame. | auto small medium large |
bgm | Whether to add background music to the generated video. | default True |
seed | The random seed to use for the generation. -1 means a random seed will be used. |
Sample prompt
The prompt behind the sample.
A professional tennis player striking a fast-moving ball on the court, powerful swing and athletic motion, bright sunlight illuminating the scene, cinematic sports shot with dynamic energy.
style: generalresolution: 1080pduration: 5generate_audio: Trueaspect_ratio: 16:9movement_amplitude: autobgm: TrueFAQ
Short answers.
How much does Vidu Q2 Text-to-video cost?
Pricing starts at $0.063 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $0.50. Usage is billed per request from your balance — no subscription.
Does Vidu Q2 Text-to-video run uncensored here?
No. Vidu applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Vidu Q2 Text-to-video for everything else it does well.
What does Vidu Q2 Text-to-video take as input?
It is a text-to-video model. Vidu Q2 Text-to-Video is an advanced AI video generation model that brings static images to life. Upload a reference image and describe the motion you want — the model generates high-quality video with smooth animation, optional audio, and cinematic quality up
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related