Gemini 2.5 Flash (batch)
Google · Multimodal · Released Jun 2025
Google's Flash model with native thinking capabilities, optimized for batch processing with a 1M token context window.
- Strengths
- Balances reasoning ability through thinking with fast inference speed, well-suited to high-throughput scenarios where latency can be traded for cost efficiency.
- Best for
- Batch jobs requiring reasoning or multi-step problem solving where processing latency is acceptable and throughput matters more than real-time response.
- Limitations
- Thinking capability adds latency compared to standard Flash models; for real-time applications or simple tasks without reasoning needs, the non-batch Gemini 2.5 Flash Lite or Gemini 3 Flash may be more suitable.
Input / 1M
$0.15
Output / 1M
$1.25
Cached input / 1M
$0.03
Context window
1.05M
Also billed
- Audio input
- $0.5 / 1M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 28 Jul 2026 | $0.15 | $1.25 | $0.03 | Imported from OpenRouter | openrouter.ai |