Qwen-Image Text-to-image Max
General-purpose image generation model that supports various art styles and is particularly good at rendering complex text.

What it does
Qwen-Image Text-to-image Max, in practice.
The flagship text-to-image generation model from Alibaba Cloud, designed to deliver state-of-the-art visual quality, exceptional prompt adherence, and rich artistic detail. Qwen-Image Max represents the pinnacle of the Qwen-Image family, capable of transforming complex text descriptions into stunning, high-resolution visuals suitable for professional and creative workflows.
- Purpose: Generate premium-quality images from natural language descriptions.
- Core Capability: Industry-leading visual fidelity with deep semantic understanding of prompts.
- Foundation: Built on Alibaba's advanced large-scale multi-modal architecture.
- Typical Output: High-resolution, photorealistic or artistic images with precise lighting, texture, and composition.
- Use Cases: Professional design, advertising creatives, concept art, marketing materials, and high-end content creation.
- Superior Visual Quality: Delivers the highest level of detail, texture, and lighting realism available in the Qwen-Image series.
Run Qwen-Image Text-to-image Max
from $0.078/imageParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | The prompt for generating the image. | |
negative_prompt | Negative prompt for the generation. | |
enable_prompt_expansion | If set to true, the prompt optimizer will be enabled. | default |
num_images | Number of images to generate. | default 1 |
size | The size of the generated image in pixels (width*height). | 1664*928 1472*1104 1328*1328 1104*1472 928*1664 |
seed | The random seed to use for the generation. -1 means a random seed will be used. | default -1 |
Sample prompt
The prompt behind the sample.
Healing-style hand-drawn poster featuring three puppies playing with a ball on lush green grass, adorned with decorative elements such as birds and stars. The main title “Come Play Ball!” is prominently displayed at the top in bold, blue cartoon font. Below it, the subtitle “Come [Show Off Your Skills]!” appears in green font. A speech bubble adds playful charm with the text: “Hehe, watch me amaze my little friends next!” At the bottom, supplementary text reads: “We get to play ball with our friends again!” The color palette centers on fresh greens and blues, accented with bright pink and yell
size: 1328*1328num_images: 1seed: -1enable_prompt_expansion: FAQ
Short answers.
How much does Qwen-Image Text-to-image Max cost?
Pricing starts at $0.078 per image at the base resolution; higher resolutions and longer durations cost more. Usage is billed per request from your balance — no subscription.
Does Qwen-Image Text-to-image Max run uncensored here?
No. Qwen applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Qwen-Image Text-to-image Max for everything else it does well.
What does Qwen-Image Text-to-image Max take as input?
It is a text-to-image model. General-purpose image generation model that supports various art styles and is particularly good at rendering complex text.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related