prices synced 2026-09-11

Gemini 3.1 Flash Lite (batch)

Google · Multimodal · Released May 2026

Compare

Google's lightweight text-only model optimized for batch processing, with a 1M token context window built on the Gemini 3.1 Flash architecture.

Strengths
Handles large context windows efficiently while maintaining fast inference speeds for text-only tasks.
Best for
Batch processing workflows where latency is flexible but throughput and context capacity matter, such as bulk document analysis or content generation jobs.
Limitations
Text-only; does not process images or multimodal input, and is superseded by Gemini 3.6 Flash (batch) as the newer batch-optimized option in Google's lineup.

Input / 1M

$0.125

Output / 1M

$0.75

Cached input / 1M

$0.0125

Context window

1.05M

Also billed

Audio input
$0.25 / 1M

Price history

Price per 1M tokens over timeOutput $0.75; Input $0.125 as of Jul 2026.$0$0.2$0.4$0.6$0.8Output on 28 Jul 2026: $0.75Out $0.75Input on 28 Jul 2026: $0.125In $0.125Jul 2026

Snapshots

Effective Input Output Cached in Note Source
28 Jul 2026 $0.125 $0.75 $0.0125 Imported from OpenRouter openrouter.ai

More from Google

Report a problem