Gemini 3.1 Flash Lite (batch)
Google · Multimodal · Released May 2026
Google's lightweight text-only model optimized for batch processing, with a 1M token context window built on the Gemini 3.1 Flash architecture.
- Strengths
- Handles large context windows efficiently while maintaining fast inference speeds for text-only tasks.
- Best for
- Batch processing workflows where latency is flexible but throughput and context capacity matter, such as bulk document analysis or content generation jobs.
- Limitations
- Text-only; does not process images or multimodal input, and is superseded by Gemini 3.6 Flash (batch) as the newer batch-optimized option in Google's lineup.
Input / 1M
$0.125
Output / 1M
$0.75
Cached input / 1M
$0.0125
Context window
1.05M
Also billed
- Audio input
- $0.25 / 1M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 28 Jul 2026 | $0.125 | $0.75 | $0.0125 | Imported from OpenRouter | openrouter.ai |