prices synced 2026-07-28

Gemini 2.5 Flash Lite Preview 09-2025

retired

Google · Multimodal · Released Sep 2025

Compare

A lightweight model in Google's Gemini 2.5 family optimized for low-latency and cost-efficient inference with a 1M token context window.

Strengths
Delivers fast token generation and high throughput, making it suitable for real-time applications where response speed matters.
Best for
Latency-sensitive workloads and high-volume inference where a smaller, faster model is preferred over maximum capability.
Limitations
Being a lite variant, it sacrifices some reasoning depth and complex task handling compared to the full Gemini 2.5 model.

Input / 1M

$0.1

Output / 1M

$0.4

Cached input / 1M

$0.01

Context window

1.05M

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $0.4 to $0.4; Input went from $0.1 to $0.1 between Jun 2026 and Jun 2026.$0$0.1$0.2$0.3$0.4Output on 11 Jun 2026: $0.4Output on 29 Jun 2026: $0.4Out $0.4Input on 11 Jun 2026: $0.1Input on 29 Jun 2026: $0.1In $0.1Jun 2026Jun 2026

Price change

30d
in decreased 0.0% out decreased 0.0%
90d
in decreased 0.0% out decreased 0.0%
1y
in decreased 0.0% out decreased 0.0%
Since launch
in decreased 0.0% out decreased 0.0%

Snapshots

Effective Input Output Cached in Note Source
29 Jun 2026 $0.1 $0.4 $0.01 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.1 $0.4 $0.01 Imported from OpenRouter openrouter.ai

More from Google

Report a problem