prices synced 2026-07-28

Gemini 3.5 Flash Lite (batch)

Google · Multimodal · Released Jul 2026

Compare

A lightweight text model from Google optimized for efficient execution of focused tasks within multi-agent agentic workflows.

Strengths
Handles subagent workloads with minimal computational overhead while maintaining a 1M token context window for processing extended documents and conversation histories.
Best for
Subagents that execute specialized subtasks as part of larger multi-agent systems, where efficiency and fast response times matter more than raw capability.
Limitations
Superseded by Gemini 3.5 Flash Lite in the batch API tier (both released 2026-07); use the standard non-batch Gemini 3.5 Flash Lite unless you specifically need batch processing for latency-insensitive workloads.

Input / 1M

$0.15

Output / 1M

$1.25

Cached input / 1M

$0.015

Context window

1.05M

Also billed

Audio input
$0.15 / 1M

Price history

Price per 1M tokens over timeOutput $1.25; Input $0.15 as of Jul 2026.$0$0.5$1.00Output on 28 Jul 2026: $1.25Out $1.25Input on 28 Jul 2026: $0.15In $0.15Jul 2026

Snapshots

Effective Input Output Cached in Note Source
28 Jul 2026 $0.15 $1.25 $0.015 Imported from OpenRouter openrouter.ai

More from Google

Report a problem