Gemini 3.1 Flash Lite
Google · Multimodal · Released May 2026
Google's lightweight text model designed for efficient execution in multi-agent and agentic workflows with a 1M token context window.
- Strengths
- Handles long contexts and is built for speed and efficiency in agent-based systems where multiple models operate together.
- Best for
- Subagent execution, multi-agent workflows, and scenarios where you need a fast, low-latency text model that can process large documents or conversation histories.
- Limitations
- Optimized for specific agentic use cases; for general-purpose text tasks or reasoning-heavy work, Gemini 3.1 Pro or Gemini 3.6 Flash may be more suitable.
Input / 1M
$0.25
Output / 1M
$1.50
Cached input / 1M
$0.025
Context window
1.05M
Also billed
- Cache write
- $0.0833 / 1M
- Audio input
- $0.5 / 1M
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in decreased 0.0% out decreased 0.0%
- 90d
- in decreased 0.0% out decreased 0.0%
- 1y
- in decreased 0.0% out decreased 0.0%
- Since launch
- in decreased 0.0% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 29 Jun 2026 | $0.25 | $1.50 | $0.025 | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.25 | $1.50 | $0.025 | Imported from OpenRouter | openrouter.ai |