prices synced 2026-07-28

Gemini 3.1 Flash Lite

Google · Multimodal · Released May 2026

Compare

Google's lightweight text model designed for efficient execution in multi-agent and agentic workflows with a 1M token context window.

Strengths
Handles long contexts and is built for speed and efficiency in agent-based systems where multiple models operate together.
Best for
Subagent execution, multi-agent workflows, and scenarios where you need a fast, low-latency text model that can process large documents or conversation histories.
Limitations
Optimized for specific agentic use cases; for general-purpose text tasks or reasoning-heavy work, Gemini 3.1 Pro or Gemini 3.6 Flash may be more suitable.

Input / 1M

$0.25

Output / 1M

$1.50

Cached input / 1M

$0.025

Context window

1.05M

Also billed

Cache write
$0.0833 / 1M
Audio input
$0.5 / 1M

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $1.50 to $1.50; Input went from $0.25 to $0.25 between Jun 2026 and Jun 2026.$0$0.5$1.00$1.50Output on 11 Jun 2026: $1.50Output on 29 Jun 2026: $1.50Out $1.50Input on 11 Jun 2026: $0.25Input on 29 Jun 2026: $0.25In $0.25Jun 2026Jun 2026

Price change

30d
in decreased 0.0% out decreased 0.0%
90d
in decreased 0.0% out decreased 0.0%
1y
in decreased 0.0% out decreased 0.0%
Since launch
in decreased 0.0% out decreased 0.0%

Snapshots

Effective Input Output Cached in Note Source
29 Jun 2026 $0.25 $1.50 $0.025 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.25 $1.50 $0.025 Imported from OpenRouter openrouter.ai

More from Google

Report a problem