Wan-2.6 Image-to-video Flash
Wan2.6 image to video flash, faster and more cost-effective generation. Intelligent shot scheduling enables multi‑camera storytelling, supports stable multi‑speaker dialogue with more natural and realistic vocal timbres.
What it does
Wan-2.6 Image-to-video Flash, in practice.
Wan2.6 image to video flash, faster and more cost-effective generation. Intelligent shot scheduling enables multi‑camera storytelling, supports stable multi‑speaker dialogue with more natural and realistic vocal timbres, and supports generation of clips up to 15 seconds in length. Alibaba WAN 2.6 Image-to-Video Flash is an advanced image-to-video model on Alibaba Cloud’s DashScope. It generates high-quality videos from images and supports output resolutions of 720p and 1080p.
- More affordable: Wan 2.6 is more streamlined and cost-effective - reducing
- One-pass A/V sync: Wan 2.6 creates a fully synchronized video
- Multilingual friendly: Wan 2.6 reliably processes like Chinese prompts
- Longer duration & more video size options: Wan 2.6 delivers up to
- Multi-shot storytelling: Generates cohesive multi-shot narratives,
- Video reference generation: Uses a reference video's appearance and
Run Wan-2.6 Image-to-video Flash
from $0.027/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
audio | Audio URL to guide generation (optional). | |
duration | The duration of the generated media in seconds. | default 5 |
enable_prompt_expansion | If set to true, the prompt optimizer will be enabled. | default True |
image | The image for generating the output. URL or base64 encoded image data. | |
negative_prompt | Negative prompt for the generation. | |
prompt | The prompt for generating the output. | |
resolution | The resolution of the generated video. | 720p 1080p |
seed | The random seed to use for the generation. -1 means a random seed will be used. | default -1 |
shot_type | Generate video in multi camera angles, only works when set enable_prompt_expansion to true. | multi single |
generate_audio | Whether to automatically add audio to the generated video. | default True |
Sample prompt
The prompt behind the sample.
A scene of urban fantasy art. A dynamic graffiti art character. A teenager, painted with spray paint, comes to life from a concrete wall. He raps at breakneck speed in an English rap while striking a classic, energetic rapper pose. The scene is set at night under a quaint urban railway bridge. Light comes from a lone streetlamp, creating a cinematic atmosphere, full of high energy and stunning detail. The audio of the video consists entirely of his rap, with no other dialogue or background noise.
seed: -1duration: 10shot_type: multiresolution: 720penable_prompt_expansion: FAQ
Short answers.
How much does Wan-2.6 Image-to-video Flash cost?
Pricing starts at $0.027 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $0.22. Usage is billed per request from your balance — no subscription.
Does Wan-2.6 Image-to-video Flash run uncensored here?
No. Qwen applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Wan-2.6 Image-to-video Flash for everything else it does well.
What does Wan-2.6 Image-to-video Flash take as input?
It is a image-to-video model. Wan2.6 image to video flash, faster and more cost-effective generation. Intelligent shot scheduling enables multi‑camera storytelling, supports stable multi‑speaker dialogue with more natural and realistic vocal timbres.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related