Compare · Text-to-video
MiniMax H3 Text-to-Video vs MiniMax H3 Fast Text-to-Video
Balanced H3 against the speed-optimised tier. The interesting part is that Fast is both cheaper and more limited in resolution — pick based on the deliverable, not the price.
Every value below is read from the model catalog at generation time — price, resolution options, duration steps, aspect ratios and which parameters each model exposes. 3 of the 11 spec rows differ between these two. “Not exposed” means the model has no API parameter for it — it does not mean the capability is absent. Audio is the clearest case: MiniMax H3 generates sound with the clip but offers no switch to control it, while some models expose one.
Specs
Side by side
| MiniMax H3 Text-to-Video | MiniMax H3 Fast Text-to-Video | |
|---|---|---|
| Vendor | MiniMax | MiniMax |
| List price | $0.057 per second | $0.066 per second |
| Category | Text-to-video | Text-to-video |
| Resolution options | 480P, 768P, 2K | 480P |
| Duration options | 4–15s (12 steps) | 5–15s (11 steps) |
| Aspect ratios | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 |
| Audio parameter exposed | not exposed | not exposed |
| Seed / reproducibility | not exposed | not exposed |
| Prompt expansion | exposed | exposed |
| Camera control parameter | not exposed | not exposed |
| Content policy | spicy route — no extra platform filter | spicy route — no extra platform filter |
MiniMax H3 Text-to-Video is the cheaper of the two at $0.057 versus $0.066 — a 1.2× difference. At five seconds that is $0.28 against $0.33.
Output
What each one actually produces
These are the real sample clips from the catalog, with the prompt that generated each. Watch both before you decide — a spec table will not tell you whether the motion reads the way you need it to.
Sample · prompt: “Cinematic medium close-up of a desert warrior wearing a high-tech dust mask and stillsuit, glowing piercing bright blue eyes (eyes of Ibad). The wind blows fine dust across their face. Warm desert sun…”
Sample · prompt: “A lone man walks through a rain-soaked neon city at midnight, reflections shimmering across the wet streets, distant headlights cutting through the mist, his coat gently swaying in the wind. The camer…”
FAQ
Short answers
Which is cheaper, MiniMax H3 Text-to-Video or MiniMax H3 Fast Text-to-Video?
MiniMax H3 Text-to-Video, at $0.057 per second. The other is listed at $0.066. These are list prices; billing on this site is not live yet.
Is MiniMax H3 Text-to-Video uncensored?
Yes — this is a MiniMax H3 route served without an extra platform refusal layer. Lawful use is your responsibility and anyone under 18 is out of scope.
Is MiniMax H3 Fast Text-to-Video uncensored?
Yes — this is a MiniMax H3 route served without an extra platform refusal layer. Lawful use is your responsibility and anyone under 18 is out of scope.
What is the maximum clip length for each?
MiniMax H3 Text-to-Video: 4–15s (12 steps). MiniMax H3 Fast Text-to-Video: 5–15s (11 steps). Longer work means multiple takes stitched together, not one longer request.
Which should I use for text-to-video?
Whichever exposes the parameter you need — the table above is the honest answer. If you are iterating, use the cheaper model until the motion is right, then re-render once. If you need a specific resolution path, check the resolution row: some tiers reach it natively and others via a super-resolution pass, which is not the same thing.