MAI-Image-2.5 Text-to-image
Microsoft's flagship text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts.

What it does
MAI-Image-2.5 Text-to-image, in practice.
MAI-Image-2.5 is Microsoft's flagship text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts. It uses a diffusion-based generative approach to progressively refine images, enabling strong alignment between the input text and the generated output. Released on June 2, 2026, it ranks among the top-performing image generation models globally.
- Photorealistic image synthesis — Generates realistic imagery with consistent visual structure, accurate lighting, depth, and texture, suitable for concept visualization and professional content creation.
- High-fidelity portraits — Produces expressive, natural-looking portraits with accurate facial structure, lighting, and skin texture.
- Accurate text rendering — Significantly improved rendering of legible text within generated images, including labels, posters, packaging, and signage.
- Visual reasoning — Reasons across objects, scene structure, lighting, scale, and spatial positioning to produce consistent outputs even from ambiguous or complex prompts.
- Product, branding & commercial design — Well suited for product imagery, marketing visuals, brand assets, and commercial creative workflows.
- Creative concept visualization — Translates abstract textual descriptions into visually coherent and imaginative outputs.
Run MAI-Image-2.5 Text-to-image
from $0.075/imageParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | Text prompt describing the image to generate. Maximum context length: 32,000 tokens. | |
size | Image dimensions in width*height format (e.g., 1024*1024, 1280*720). Minimum 768. The product of width × height must not exceed 1,049,088. | default 1024*1024 |
Sample prompt
The prompt behind the sample.
A futuristic fighter jet weaves through the clouds.
size: 1280*768enable_base64_output: enable_sync_mode: FAQ
Short answers.
How much does MAI-Image-2.5 Text-to-image cost?
Pricing starts at $0.075 per image at the base resolution; higher resolutions and longer durations cost more. Usage is billed per request from your balance — no subscription.
Does MAI-Image-2.5 Text-to-image run uncensored here?
No. Microsoft applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists MAI-Image-2.5 Text-to-image for everything else it does well.
What does MAI-Image-2.5 Text-to-image take as input?
It is a text-to-image model. Microsoft's flagship text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related