Grok Imagine Image Text-to-Image
xAI Grok Imagine generates images from natural-language prompts at 1K or 2K resolution, with 14 aspect ratios.

What it does
Grok Imagine Image Text-to-Image, in practice.
Run Grok Imagine Image Text-to-Image
from $0.030/imageParameters
What you can set.
| Parameter | What it does | Options |
|---|---|---|
prompt | Natural-language description of the image to generate. | |
num_images | Number of images to generate. Each image is billed separately. | 1 2 3 4 |
aspect_ratio | Aspect ratio of the generated image. | 1:1 3:4 4:3 9:16 16:9 2:3 3:2 9:19.5 19.5:9 9:20 20:9 1:2 |
resolution | Output resolution. 1k = 1024x1024, 2k = 2048x2048. | 1k 2k |
Sample prompt
The prompt behind the sample.
Ancient futuristic city carved into towering desert cliffs, monumental architecture, vast dunes surrounding the city, warm golden tones, mysterious atmosphere, cinematic sci-fi worldbuilding, ultra detailed, epic scale, volumetric sunlight, Dune aesthetic
num_images: 1aspect_ratio: 16:9resolution: 1kenable_base64_output: FAQ
Short answers.
How much does Grok Imagine Image Text-to-Image cost?
Pricing starts at $0.030 per image at the base resolution; higher resolutions and longer durations cost more. Usage is billed per request from your balance — no subscription.
Does Grok Imagine Image Text-to-Image run uncensored here?
No. xAI applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Grok Imagine Image Text-to-Image for everything else it does well.
What does Grok Imagine Image Text-to-Image take as input?
It is a text-to-image model. xAI Grok Imagine generates images from natural-language prompts at 1K or 2K resolution, with 14 aspect ratios.
How do I use it?
Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.
Related