MiniMax H3 Text-to-Video
MiniMax H3 text-to-video: generate a cinematic video from a text prompt. Supports 2K, 5-15s., and 16:9/9:16/1:1/adaptive aspect ratios.
What it does
MiniMax H3 Text-to-Video, in practice.
MiniMax H3 Text-to-Video is a state-of-the-art AI video generation model that creates cinematic, high-fidelity videos directly from a text prompt. With crisp detail up to 2K, smooth natural motion, and flexible aspect ratios, it turns a single description into a polished clip.
- High resolution output Generate videos in 2K quality.
- Cinematic motion Fluid camera work and lifelike movement from a plain-text description.
- Flexible aspect ratios 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 or adaptive to fit any platform.
- Selectable duration Produce 4–15s clips.
Run MiniMax H3 Text-to-Video
from $0.057/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | The text prompt describing the video to generate. | |
resolution | The resolution of the generated video. | 480P 768P 2K |
duration | The duration of the generated video in seconds. | 4 5 6 7 8 9 10 11 12 13 14 15 |
ratio | The aspect ratio of the generated video. | 21:9 16:9 4:3 1:1 3:4 9:16 |
prompt_expansion | Whether to expand the prompt for better results. | default |
Sample prompt
The prompt behind the sample.
Cinematic medium close-up of a desert warrior wearing a high-tech dust mask and stillsuit, glowing piercing bright blue eyes (eyes of Ibad). The wind blows fine dust across their face. Warm desert sunlight casting sharp side shadows. Cinematic focus, shallow depth of field, slow-motion film grain, photorealistic, Dune style, 8k.
resolution: 2Kduration: 8ratio: 16:9FAQ
Short answers.
How much does MiniMax H3 Text-to-Video cost?
Pricing starts at $0.057 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $0.46. Usage is billed per request from your balance — no subscription.
Is this the uncensored route?
Yes. This is the MiniMax H3 route served here without an extra platform refusal layer on top of the model. Lawful prompts and outputs are your responsibility; anyone under 18 is out of scope.
What does MiniMax H3 Text-to-Video take as input?
It is a text-to-video model. MiniMax H3 text-to-video: generate a cinematic video from a text prompt. Supports 2K, 5-15s., and 16:9/9:16/1:1/adaptive aspect ratios.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related