prices synced 2026-09-11

Gemini 3.5 Flash-Lite

Google · Multimodal · Released Jul 2026

Compare

A lightweight text-only model from Google with a 1M token context window, released in mid-2026 as a faster, more efficient alternative to the standard Flash tier.

Strengths
Handles large text workloads with lower latency and reduced resource consumption compared to the full Flash model.
Best for
Processing text at scale — summarization, classification, bulk analysis, and other tasks where speed and efficiency matter more than advanced reasoning.
Limitations
Text-only; lacks the multimodal support (image, video, audio) available in later Flash models and cannot handle the full range of complex reasoning tasks that larger models support.

Input / 1M

$0.3

Output / 1M

$2.50

Cached input / 1M

$0.03

Context window

1.05M

Also billed

Cache write
$0.0833 / 1M
Audio input
$0.3 / 1M

Price history

Price per 1M tokens over timeOutput $2.50; Input $0.3 as of Jul 2026.$0$1.00$2.00Output on 21 Jul 2026: $2.50Out $2.50Input on 21 Jul 2026: $0.3In $0.3Jul 2026

Snapshots

Effective Input Output Cached in Note Source
21 Jul 2026 $0.3 $2.50 $0.03 Imported from OpenRouter openrouter.ai

More from Google

Report a problem