BLACKFORESTLABS FLUX 3 Text-to-Video
Generate a video (with audio) from a text prompt using FLUX 3.
What it does
BLACKFORESTLABS FLUX 3 Text-to-Video, in practice.
FLUX 3 Text-to-Video generates cinematic videos with synchronized audio directly from a text prompt. Powered by Black Forest Labs' FLUX 3 architecture, it delivers high-fidelity motion, strong prompt adherence, and native audio generation — no post-processing required.
- Native audio generation Produce video and audio together in a single request — dialogue, ambient sound, and music matched to the scene.
- Explicit duration required Duration must be set explicitly to validate keyframe positions.
- Auto aspect ratio Use `auto` to let the model pick the best fit, or specify any ratio from ultra-wide to portrait.
- High resolution output Generate videos in 720p or 1080p.
- Adjustable safety Fine-tune the safety tolerance from strict (0) to permissive (4).
Run BLACKFORESTLABS FLUX 3 Text-to-Video
from $0.26/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | Text description of the video to generate. | |
aspect_ratio | Aspect ratio of the generated video. auto lets the model choose. | auto 21:9 2:1 16:9 4:3 1:1 3:4 9:16 |
resolution | Resolution of the generated video. | 720p 1080p |
duration | Duration of the generated video in seconds. An explicit duration is required to place the end frame. | 5 6 7 8 9 10 11 12 13 14 15 16 |
generate_audio | Whether to generate audio for the video. | default True |
safety_tolerance | Safety tolerance level. 0 is the strictest and 4 is the most permissive. | default 2 |
seed | Random seed for reproducibility. |
Sample prompt
The prompt behind the sample.
A powerful character unleashing an overwhelming energy force, pushing an opponent across the battlefield while space and time distort around them. A massive shockwave expands through the air, bending gravity, twisting light, and creating cracks in the fabric of reality. The opponent is thrown into the distance surrounded by swirling energy fragments. Cinematic action scene, epic scale, dynamic composition, dramatic lighting, ultra realistic, Hollywood sci-fi movie style, 8K, masterpiece.
aspect_ratio: 16:9resolution: 720pduration: 5generate_audio: Truesafety_tolerance: 2FAQ
Short answers.
How much does BLACKFORESTLABS FLUX 3 Text-to-Video cost?
Pricing starts at $0.26 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $2.04. Usage is billed per request from your balance — no subscription.
Does BLACKFORESTLABS FLUX 3 Text-to-Video run uncensored here?
No. Black Forest Labs applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists BLACKFORESTLABS FLUX 3 Text-to-Video for everything else it does well.
What does BLACKFORESTLABS FLUX 3 Text-to-Video take as input?
It is a text-to-video model. Generate a video (with audio) from a text prompt using FLUX 3.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related