Gemini 3.5 Flash-Lite
Google · Multimodal · Released Jul 2026
A lightweight text-only model from Google with a 1M token context window, released in mid-2026 as a faster, more efficient alternative to the standard Flash tier.
- Strengths
- Handles large text workloads with lower latency and reduced resource consumption compared to the full Flash model.
- Best for
- Processing text at scale — summarization, classification, bulk analysis, and other tasks where speed and efficiency matter more than advanced reasoning.
- Limitations
- Text-only; lacks the multimodal support (image, video, audio) available in later Flash models and cannot handle the full range of complex reasoning tasks that larger models support.
Input / 1M
$0.3
Output / 1M
$2.50
Cached input / 1M
$0.03
Context window
1.05M
Also billed
- Cache write
- $0.0833 / 1M
- Audio input
- $0.3 / 1M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 21 Jul 2026 | $0.3 | $2.50 | $0.03 | Imported from OpenRouter | openrouter.ai |