prices synced 2026-09-12

Gemini 2.5 Flash (batch)

Google · Multimodal · Released Jun 2025

Compare

Google's Flash model with native thinking capabilities, optimized for batch processing with a 1M token context window.

Strengths
Balances reasoning ability through thinking with fast inference speed, well-suited to high-throughput scenarios where latency can be traded for cost efficiency.
Best for
Batch jobs requiring reasoning or multi-step problem solving where processing latency is acceptable and throughput matters more than real-time response.
Limitations
Thinking capability adds latency compared to standard Flash models; for real-time applications or simple tasks without reasoning needs, the non-batch Gemini 2.5 Flash Lite or Gemini 3 Flash may be more suitable.

Input / 1M

$0.15

Output / 1M

$1.25

Cached input / 1M

$0.03

Context window

1.05M

Also billed

Audio input
$0.5 / 1M

Price history

Price per 1M tokens over timeOutput $1.25; Input $0.15 as of Jul 2026.$0$0.5$1.00Output on 28 Jul 2026: $1.25Out $1.25Input on 28 Jul 2026: $0.15In $0.15Jul 2026

Snapshots

Effective Input Output Cached in Note Source
28 Jul 2026 $0.15 $1.25 $0.03 Imported from OpenRouter openrouter.ai

More from Google

Report a problem