MAI-Image-2.5-Flash Text-to-image
Microsoft's fast, cost-optimized text-to-image generation model, creating high-quality images at lower cost using the same diffusion-based architecture as MAI-Image-2.5.

What it does
MAI-Image-2.5-Flash Text-to-image, in practice.
MAI-Image-2.5-Flash is Microsoft's fast, cost-optimized text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts at significantly lower cost than the standard MAI-Image-2.5. It uses the same diffusion-based generative approach, enabling strong alignment between the input text and the generated output, while being optimized for speed and throughput. Released on June 2, 2026.
- Photorealistic image synthesis — Generates realistic imagery with consistent visual structure, accurate lighting, depth, and texture, suitable for concept visualization and professional content creation.
- High-fidelity portraits — Produces expressive, natural-looking portraits with accurate facial structure, lighting, and skin texture.
- Accurate text rendering — Improved rendering of legible text within generated images, including labels, posters, packaging, and signage.
- Visual reasoning — Reasons across objects, scene structure, lighting, scale, and spatial positioning to produce consistent outputs even from ambiguous or complex prompts.
- Product, branding & commercial design — Well suited for product imagery, marketing visuals, brand assets, and commercial creative workflows.
- Creative concept visualization — Translates abstract textual descriptions into visually coherent and imaginative outputs.
Run MAI-Image-2.5-Flash Text-to-image
from $0.045/imageParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | Text prompt describing the image to generate. Maximum context length: 32,000 tokens. | |
size | Image dimensions in width*height format (e.g., 1024*1024, 1280*720). Minimum 768. The product of width × height must not exceed 1,049,088. | default 1024*1024 |
Sample prompt
The prompt behind the sample.
Generate a Matrix-themed poster
size: 768*1280enable_base64_output: enable_sync_mode: FAQ
Short answers.
How much does MAI-Image-2.5-Flash Text-to-image cost?
Pricing starts at $0.045 per image at the base resolution; higher resolutions and longer durations cost more. Usage is billed per request from your balance — no subscription.
Does MAI-Image-2.5-Flash Text-to-image run uncensored here?
No. Microsoft applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists MAI-Image-2.5-Flash Text-to-image for everything else it does well.
What does MAI-Image-2.5-Flash Text-to-image take as input?
It is a text-to-image model. Microsoft's fast, cost-optimized text-to-image generation model, creating high-quality images at lower cost using the same diffusion-based architecture as MAI-Image-2.5.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related