prices synced 2026-09-11

Gemma 3n 4B

Google · Released May 2025

Compare

Google's 4-billion-parameter open-weight text model designed for efficient local deployment or serverless inference.

Strengths
Fast inference and low memory footprint across devices, making it suitable for resource-constrained environments and edge deployment.
Best for
Applications requiring quick responses on limited hardware, such as mobile inference, embedded systems, or high-throughput serverless workloads.
Limitations
At the smallest tier of Google's Gemma 3 lineup, it trades reasoning depth and context understanding for speed; larger Gemma 3 variants (12B and 27B) handle more complex tasks better.

Input / 1M

$0.06

Output / 1M

$0.12

Cached input / 1M

Context window

32K

Price history

Price per 1M tokens over timeOutput $0.12; Input $0.06 as of Jun 2026.$0$0.05$0.1Output on 11 Jun 2026: $0.12Out $0.12Input on 11 Jun 2026: $0.06In $0.06Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.06 $0.12 Imported from OpenRouter openrouter.ai

More from Google

Report a problem