minimaxiH3 Get access
MicrosoftText-to-imageVendor content policy applies

MAI-Image-2.5 Text-to-image

Microsoft's flagship text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts.

$0.075per imageStarting price at the base resolution and quality tier.
$0.75Ten outputs at the base tier. Larger sizes and premium quality tiers cost more.
Pay per useBilled per request from your balance. No subscription, no minimum.
MAI-Image-2.5 Text-to-image sample output
Sample output · prompt: “A futuristic fighter jet weaves through the clouds.”

What it does

MAI-Image-2.5 Text-to-image, in practice.

MAI-Image-2.5 is Microsoft's flagship text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts. It uses a diffusion-based generative approach to progressively refine images, enabling strong alignment between the input text and the generated output. Released on June 2, 2026, it ranks among the top-performing image generation models globally.

  • Photorealistic image synthesis — Generates realistic imagery with consistent visual structure, accurate lighting, depth, and texture, suitable for concept visualization and professional content creation.
  • High-fidelity portraits — Produces expressive, natural-looking portraits with accurate facial structure, lighting, and skin texture.
  • Accurate text rendering — Significantly improved rendering of legible text within generated images, including labels, posters, packaging, and signage.
  • Visual reasoning — Reasons across objects, scene structure, lighting, scale, and spatial positioning to produce consistent outputs even from ambiguous or complex prompts.
  • Product, branding & commercial design — Well suited for product imagery, marketing visuals, brand assets, and commercial creative workflows.
  • Creative concept visualization — Translates abstract textual descriptions into visually coherent and imaginative outputs.

Run MAI-Image-2.5 Text-to-image

from $0.075/image
Invite code opens Chat with this model loaded. No code yet? Join the waitlist — we count which models people ask for.

Parameters

What you can set.

ParameterWhat it doesOptions
promptText prompt describing the image to generate. Maximum context length: 32,000 tokens.
sizeImage dimensions in width*height format (e.g., 1024*1024, 1280*720). Minimum 768. The product of width × height must not exceed 1,049,088.default 1024*1024

Sample prompt

The prompt behind the sample.

A futuristic fighter jet weaves through the clouds.
size: 1280*768enable_base64_output: enable_sync_mode:

FAQ

Short answers.

How much does MAI-Image-2.5 Text-to-image cost?

Pricing starts at $0.075 per image at the base resolution; higher resolutions and longer durations cost more. Usage is billed per request from your balance — no subscription.

Does MAI-Image-2.5 Text-to-image run uncensored here?

No. Microsoft applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists MAI-Image-2.5 Text-to-image for everything else it does well.

What does MAI-Image-2.5 Text-to-image take as input?

It is a text-to-image model. Microsoft's flagship text-to-image generation model, designed to create high-quality, visually rich images from natural language prompts.

How do I use it?

Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.

Related

Models people compare with this one.

Run MAI-Image-2.5 Text-to-image.

Invite code opens Chat with the model loaded. No code — join the waitlist and we will count the request.

Get access Browse models