Wan-2.7 Text-to-video
Generates videos from text prompts with multi-shot narrative, audio generation, and sound-image synchronization.
What it does
Wan-2.7 Text-to-video, in practice.
Alibaba WAN 2.7 Text-to-Video generates videos from text prompts with built-in audio generation, multi-shot narrative control, and sound-image synchronization.
- Multi-shot storytelling: Describe scene-by-scene shots in the prompt, and the model generates a coherent multi-shot video with natural transitions.
- Built-in audio: Generates matching sound effects, music, and ambient audio automatically based on the prompt content.
- Flexible framing: Supports five aspect ratios (16:9, 9:16, 1:1, 4:3, 3:4) at 720P or 1080P resolution.
- Up to 15 seconds: Generate videos from 2 to 15 seconds in a single request.
- Super Resolution options: Choose `1080P-SR` or `1440P-SR` when you need a sharper final video with cleaner edges and improved texture detail.
- Short-form video creators producing social media clips, ads, and story reels.
Run Wan-2.7 Text-to-video
from $0.15/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | Text prompt describing the video content. Maximum length is 5000 characters. | |
negative_prompt | Text describing elements to exclude from the video. Maximum length is 500 characters. | |
audio | URL of an audio file to use as the video soundtrack. Supported formats: wav, mp3. Duration: 2–30 seconds. Maximum file size: 15 MB. The alias `audio_url` is also accepted. | |
resolution | Output video resolution. Higher resolution increases cost. | 720P 1080P |
ratio | Aspect ratio of the generated video. | 16:9 9:16 1:1 4:3 3:4 |
duration | Video duration in seconds. Longer duration increases cost. | default 5 |
prompt_extend | Whether to use AI to enhance the prompt for better video quality. Increases generation time. | default True |
seed | Random seed for video generation. Range: 0 to 2147483647. Use -1 for a random seed. | default -1 |
Sample prompt
The prompt behind the sample.
A supercar explodes out of a long dark tunnel at extreme speed, sunlight blindingly illuminating the road ahead, motion blur streaking past the tunnel walls, dust and light rays bursting into the scene, cinematic blockbuster style, ultra-realistic.
resolution: 1080Pratio: 16:9duration: 5prompt_extend: Trueseed: -1FAQ
Short answers.
How much does Wan-2.7 Text-to-video cost?
Pricing starts at $0.15 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $1.20. Usage is billed per request from your balance — no subscription.
Does Wan-2.7 Text-to-video run uncensored here?
No. Qwen applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Wan-2.7 Text-to-video for everything else it does well.
What does Wan-2.7 Text-to-video take as input?
It is a text-to-video model. Generates videos from text prompts with multi-shot narrative, audio generation, and sound-image synchronization.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related