minimaxiH3 Get access
GoogleText-to-imageVendor content policy applies

Nano Banana Pro Text-to-image

Nano Banana Pro is the next-generation Nano Banana image model, delivering sharper detail, richer color control, and faster diffusion for production-ready visuals.

$0.21per imageStarting price at the base resolution and quality tier.
$2.10Ten outputs at the base tier. Larger sizes and premium quality tiers cost more.
Pay per useBilled per request from your balance. No subscription, no minimum.
Nano Banana Pro Text-to-image sample output
Sample output · prompt: “Create a modern and minimalist magazine cover for April 2025, themed around **quiet city living**. The background features soft neutral tones, with a subtle gradient evoking calm and tranquility. Centered is a young pers…”

What it does

Nano Banana Pro Text-to-image, in practice.

Nano Banana Pro, officially designated as Gemini 3 Pro Image, represents the next generation in Google's series of highly-capable, natively multimodal models. It is designed for professional asset production, integrating the advanced reasoning capabilities of the Gemini 3 Pro foundation model with a sophisticated image generation engine. The primary goal of Nano Banana Pro is to provide users with studio-quality precision and control, enabling the creation of complex, high-fidelity visuals from textual and image-based prompts. Its core contribution lies in its ability to understand and execute intricate instructions, maintain character and scene consistency, and render legible text directly

  • Superior Text Rendering: The model excels at generating images that contain clear, accurate, and stylistically coherent text, making it ideal for creating posters, diagrams, and marketing materials.
  • Advanced Creative Controls: Users can exercise fine-grained control over image outputs, including camera angles, lighting transformations (e.g., day to night), color grading, depth of field, and localized editing.
  • High-Fidelity Consistency: It can maintain the consistency of up to 14 input images and blend up to 5 distinct characters seamlessly into complex compositions, ensuring visual coherence across a series of generated images.
  • Deep Real-World Knowledge: Built on Gemini 3 Pro, the model leverages a vast understanding of the world to generate contextually rich and factually grounded visuals, from detailed infographics to historically accurate scenes.
  • Multilingual Capabilities: The model can accurately render and translate text across multiple languages within an image, facilitating the localization of visual content.
  • Complex Composition from Multiple Inputs: Nano Banana Pro can synthesize elements from multiple source images and text prompts to create a single, cohesive scene, enabling complex creative concepts.

Run Nano Banana Pro Text-to-image

from $0.21/image
aspect_ratio
resolution
output_format
Invite code opens Chat with this model loaded. No code yet? Join the waitlist — we count which models people ask for.

Parameters

What you can set.

ParameterWhat it doesOptions
aspect_ratioThe aspect ratio of the generated media.1:1 3:2 2:3 3:4 4:3 4:5 5:4 9:16 16:9 21:9
enable_web_searchIf enabled, the model will use web search to ground the generation with real-time information.default
promptThe positive prompt for the generation.
resolutionThe resolution of the output image.1k 2k 4k
output_formatThe format of the output image.default png jpeg
media_resolutionControls how input media is processed. LOW reduces tokens per image/video, possibly losing detail but allowing longer videos in context. Supported values: HIGH, MEDIUM, LOW.default low medium high
seedRandom seed. Does not guarantee determinism but may improve repeatability. -1 means a random seed will be used.default -1
top_pProbability threshold for top-p samplingdefault 0.95
temperatureCreativity allowed in the responses. Best results at default 1.0. Lower values may impact reasoning.default 1

Sample prompt

The prompt behind the sample.

Create a modern and minimalist magazine cover for April 2025, themed around **quiet city living**. The background features soft neutral tones, with a subtle gradient evoking calm and tranquility. Centered is a young person sitting in a peaceful, minimalist apartment, with clean lines and a cozy, simple aesthetic. They wear a relaxed, casual outfit—comfortable sweater and pants—sitting in a calm, introspective pose. The setting includes a few plants and soft, natural lighting streaming through the window. Include magazine headlines in stylish, simple fonts, with neutral color accents: • Top le
enable_base64_output: enable_sync_mode: output_format: defaultresolution: 2kaspect_ratio: 3:4

FAQ

Short answers.

How much does Nano Banana Pro Text-to-image cost?

Pricing starts at $0.21 per image at the base resolution; higher resolutions and longer durations cost more. Usage is billed per request from your balance — no subscription.

Does Nano Banana Pro Text-to-image run uncensored here?

No. Google applies its own content policy to this model regardless of where it is called from. For a route without an extra platform filter use Uncensored MiniMax H3; this page lists Nano Banana Pro Text-to-image for everything else it does well.

What does Nano Banana Pro Text-to-image take as input?

It is a text-to-image model. Nano Banana Pro is the next-generation Nano Banana image model, delivering sharper detail, richer color control, and faster diffusion for production-ready visuals.

How do I use it?

Enter an invite code to open Chat with the model loaded, or join the waitlist. We open seats in batches and track which models are requested most.

Related

Models people compare with this one.

Run Nano Banana Pro Text-to-image.

Invite code opens Chat with the model loaded. No code — join the waitlist and we will count the request.

Get access Browse models