Gemini Omni Flash Text-to-Video Developer
Gemini Omni Flash is Google's multimodal video generation model. This text-to-video variant generates high-quality cinematic videos from text prompts with support for multiple resolutions, aspect ratios, and controllable duration.
What it does
Gemini Omni Flash Text-to-Video Developer, in practice.
Model ID: google/gemini-omni-flash/text-to-video-developer Gemini Omni is Google's multimodal video generation model designed to create high-quality video content from diverse input types. This variant accepts a text prompt only, making it ideal for pure creative generation where you describe the scene entirely through language.
- Rich prompt understanding — Describe subjects, actions, camera movements, lighting, mood, and style in a single prompt of up to 20,000 characters.
- Multi-resolution output — Generate at 720p, 1080p, or 4K.
- Flexible aspect ratios — 16:9 landscape or 9:16 portrait.
- Controllable duration — 4, 6, 8, or 10 seconds per generation.
- Reproducible results — Set a fixed seed to reproduce or iterate on a specific generation.
Run Gemini Omni Flash Text-to-Video Developer
from $0.17/secParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | Text prompt for generation. Describes the target content, style, camera language, or character actions. Maximum 20,000 characters. | |
duration | The duration of the generated video in seconds. | 4 6 8 10 |
aspect_ratio | The aspect ratio of the generated video. | 16:9 9:16 |
resolution | The resolution of the generated video. | 720p 1080p 4k |
seed | Random seed for reproducibility. Use -1 to use a random seed. | default -1 |
Sample prompt
The prompt behind the sample.
System / Style: Ultra-realistic 3D stop-motion texture, macro photography, shallow depth of field, 4K resolution. Atmospheric and emotive sound design. Visual & Action Sequence: A realistic, weathered porcelain white deer statue stands frozen in the center of a dark, damp mossy forest. The camera is locked in a tight macro shot focusing on the deer's eye. Suddenly, a single drop of glowing, golden liquid honey drips from a branch above, landing perfectly into the deer's porcelain eye. The camera slowly zooms out. Where the honey touches, the cold porcelain instantly cracks and transforms into
duration: 10aspect_ratio: 16:9resolution: 720pseed: -1FAQ
Short answers.
How much does Gemini Omni Flash Text-to-Video Developer cost?
Pricing starts at $0.17 per second of video at the base resolution; higher resolutions and longer durations cost more. An 8-second clip at the base tier is about $1.34. Usage is billed per request from your balance — no subscription.
Does Gemini Omni Flash Text-to-Video Developer run uncensored here?
No. Google applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Gemini Omni Flash Text-to-Video Developer for everything else it does well.
What does Gemini Omni Flash Text-to-Video Developer take as input?
It is a text-to-video model. Gemini Omni Flash is Google's multimodal video generation model. This text-to-video variant generates high-quality cinematic videos from text prompts with support for multiple resolutions, aspect ratios, and controllable duration.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related