Gemini 3.5 Flash Lite (batch)
Google · Multimodal · Released Jul 2026
A lightweight text model from Google optimized for efficient execution of focused tasks within multi-agent agentic workflows.
- Strengths
- Handles subagent workloads with minimal computational overhead while maintaining a 1M token context window for processing extended documents and conversation histories.
- Best for
- Subagents that execute specialized subtasks as part of larger multi-agent systems, where efficiency and fast response times matter more than raw capability.
- Limitations
- Superseded by Gemini 3.5 Flash Lite in the batch API tier (both released 2026-07); use the standard non-batch Gemini 3.5 Flash Lite unless you specifically need batch processing for latency-insensitive workloads.
Input / 1M
$0.15
Output / 1M
$1.25
Cached input / 1M
$0.015
Context window
1.05M
Also billed
- Audio input
- $0.15 / 1M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 28 Jul 2026 | $0.15 | $1.25 | $0.015 | Imported from OpenRouter | openrouter.ai |