prices synced 2026-07-28

Gemini 2.5 Flash Lite

Google · Multimodal · Released Jul 2025

Compare

Google's lightweight multimodal model with a 1M token context window, released in 2025 as a faster alternative to the Gemini 2.5 Flash line.

Strengths
Handles large context windows efficiently and processes images alongside text without specialized preprocessing.
Best for
High-volume inference tasks and workflows where speed and long-context reasoning matter more than frontier capability.
Limitations
Smaller model size and earlier release date mean it trails Gemini 3 Flash and later releases in reasoning depth and instruction-following across complex tasks.

Input / 1M

$0.1

Output / 1M

$0.4

Cached input / 1M

$0.01

Context window

1.05M

Also billed

Cache write
$0.0833 / 1M
Audio input
$0.3 / 1M

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $0.4 to $0.4; Input went from $0.1 to $0.1 between Jun 2026 and Jun 2026.$0$0.1$0.2$0.3$0.4Output on 11 Jun 2026: $0.4Output on 29 Jun 2026: $0.4Out $0.4Input on 11 Jun 2026: $0.1Input on 29 Jun 2026: $0.1In $0.1Jun 2026Jun 2026

Price change

30d
in decreased 0.0% out decreased 0.0%
90d
in decreased 0.0% out decreased 0.0%
1y
in decreased 0.0% out decreased 0.0%
Since launch
in decreased 0.0% out decreased 0.0%

Snapshots

Effective Input Output Cached in Note Source
29 Jun 2026 $0.1 $0.4 $0.01 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.1 $0.4 $0.01 Imported from OpenRouter openrouter.ai

More from Google

Report a problem