Nano Banana 2 (Gemini 3.1 Flash Image)
Google · Image generation · Released Jun 2026
Google's multimodal model for processing text and images, built on the Gemini 3.1 Flash architecture with a 131K token context window.
- Strengths
- Handles both text and image inputs in a single pass, balancing inference speed with multimodal capability on the Gemini 3.1 generation.
- Best for
- Tasks requiring image understanding alongside text processing where the larger context window and multimodal support of newer Flash models are unnecessary.
- Limitations
- Superseded by later multimodal Flash releases (3.6 and 3.7) that offer 1M token context windows; text-only workloads are better served by the lightweight Gemini 3.1 Flash Lite.
Input / 1M
$0.5
Output / 1M
$3.00
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 18 Jun 2026 | $0.5 | $3.00 | — | Imported from OpenRouter | openrouter.ai |