Gemini-3-Pro-Image (Nano Banana Pro) is a high-performance image generation and editing model built on Gemini 3 Pro. It delivers enhanced multimodal understanding and real-world semantic reasoning, enabling fast creation of well-structured visual content such as infographics, product sketches, and multi-subject scenes. It can also leverage real-time knowledge through Search grounding. The model excels in text rendering, consistent multi-image blending, and identity preservation, while offering fine-grained creative controls like localized edits, lighting and focus adjustments, camera transformations, and flexible aspect ratios. It’s ideal for rapid design, concept previews, product visualization, and everyday image generation workflows.
Pricing
Input price: $2/M, Text output: $12/M, Image output: $120/M (approximately $0.134 per 1K & 2K image, and $0.24 per 4K image)
Input Modalities
- Text
- Vision
Output Modalities
- Text
- Image
Context length
- 65.5K tokens
Max output
- 32.8K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Multimodal output
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
VertexAI gemini-3-pro-image
Pricing$2$12
Input Image$2/M tokens
Output Image$120/M tokens
Context1M
Max output65K
Latency41.9S
Throughput47.9TPS
Uptime
80.31% uptime 2 days ago
89.95% uptime yesterday
90.47% uptime today
Google AI Studio gemini-3-pro-image
Pricing$2$12
Input Image$2/M tokens
Output Image$120/M tokens
Context1M
Max output65K
Latency63.2S
Throughput26.5TPS
Uptime
58.78% uptime 2 days ago
79.91% uptime yesterday
28.31% uptime today
Performance for gemini-3-pro-image
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
