Gemini-3-Pro-Image (Nano Banana Pro) is a high-performance image generation and editing model built on Gemini 3 Pro. It delivers enhanced multimodal understanding and real-world semantic reasoning, enabling fast creation of well-structured visual content such as infographics, product sketches, and multi-subject scenes. It can also leverage real-time knowledge through Search grounding. The model excels in text rendering, consistent multi-image blending, and identity preservation, while offering fine-grained creative controls like localized edits, lighting and focus adjustments, camera transformations, and flexible aspect ratios. It’s ideal for rapid design, concept previews, product visualization, and everyday image generation workflows.
Pricing
Input price: $2/M, Text output: $12/M, Image output: $120/M (approximately $0.134 per 1K & 2K image, and $0.24 per 4K image)
Input Modalities
- Text
- Vision
Output Modalities
- Text
- Image
Context length
- 65.5K tokens
Max output
- 32.8K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Multimodal output
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
VertexAI gemini-3-pro-image
Pricing$2.000$12.000
Input Image$2/M tokens
Output Image$120/M tokens
Context1M
Max output65K
Latency40.0S
Throughput53.2TPS
Uptime
98.96% uptime 2 days ago
94.94% uptime yesterday
98.11% uptime today
Google AI Studio gemini-3-pro-image
Pricing$2.000$12.000
Input Image$2/M tokens
Output Image$120/M tokens
Context1M
Max output65K
Latency23.9S
Throughput63.2TPS
Uptime
81.59% uptime 2 days ago
68.08% uptime yesterday
66.34% uptime today
Performance for gemini-3-pro-image
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
