Gemini 2.5 Flash Lite
Google · Multimodal · Released Jul 2025
Google's lightweight multimodal model with a 1M token context window, released in 2025 as a faster alternative to the Gemini 2.5 Flash line.
- Strengths
- Handles large context windows efficiently and processes images alongside text without specialized preprocessing.
- Best for
- High-volume inference tasks and workflows where speed and long-context reasoning matter more than frontier capability.
- Limitations
- Smaller model size and earlier release date mean it trails Gemini 3 Flash and later releases in reasoning depth and instruction-following across complex tasks.
Input / 1M
$0.1
Output / 1M
$0.4
Cached input / 1M
$0.01
Context window
1.05M
Also billed
- Cache write
- $0.0833 / 1M
- Audio input
- $0.3 / 1M
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in decreased 0.0% out decreased 0.0%
- 90d
- in decreased 0.0% out decreased 0.0%
- 1y
- in decreased 0.0% out decreased 0.0%
- Since launch
- in decreased 0.0% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 29 Jun 2026 | $0.1 | $0.4 | $0.01 | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.1 | $0.4 | $0.01 | Imported from OpenRouter | openrouter.ai |