Gemma 3n 4B
Google · Released May 2025
Google's 4-billion-parameter open-weight text model designed for efficient local deployment or serverless inference.
- Strengths
- Fast inference and low memory footprint across devices, making it suitable for resource-constrained environments and edge deployment.
- Best for
- Applications requiring quick responses on limited hardware, such as mobile inference, embedded systems, or high-throughput serverless workloads.
- Limitations
- At the smallest tier of Google's Gemma 3 lineup, it trades reasoning depth and context understanding for speed; larger Gemma 3 variants (12B and 27B) handle more complex tasks better.
Input / 1M
$0.06
Output / 1M
$0.12
Cached input / 1M
—
Context window
32K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.06 | $0.12 | — | Imported from OpenRouter | openrouter.ai |