Llama 3.2 1B Instruct
Meta · Released Sep 2024
A 1-billion-parameter instruction-tuned language model from Meta with a 60,000-token context window, positioned as the smallest in the Llama 3.2 generation.
- Strengths
- Llama 3.2 1B trades capability for speed and efficiency, making it suitable for latency-sensitive tasks and edge deployments where inference costs matter.
- Best for
- Mobile applications, real-time systems, and workloads where response speed and computational footprint outweigh the need for complex reasoning or world knowledge.
- Limitations
- At 1 billion parameters, this model has limited reasoning ability, smaller knowledge base, and lower accuracy compared to larger Llama 3.2 variants like the 3B or 11B versions.
Input / 1M
$0.027
Output / 1M
$0.201
Cached input / 1M
—
Context window
60K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.027 | $0.201 | — | Imported from OpenRouter | openrouter.ai |