prices synced 2026-09-11

Llama 3.2 1B Instruct

Meta · Released Sep 2024

Compare

A 1-billion-parameter instruction-tuned language model from Meta with a 60,000-token context window, positioned as the smallest in the Llama 3.2 generation.

Strengths
Llama 3.2 1B trades capability for speed and efficiency, making it suitable for latency-sensitive tasks and edge deployments where inference costs matter.
Best for
Mobile applications, real-time systems, and workloads where response speed and computational footprint outweigh the need for complex reasoning or world knowledge.
Limitations
At 1 billion parameters, this model has limited reasoning ability, smaller knowledge base, and lower accuracy compared to larger Llama 3.2 variants like the 3B or 11B versions.

Input / 1M

$0.027

Output / 1M

$0.201

Cached input / 1M

Context window

60K

Price history

Price per 1M tokens over timeOutput $0.201; Input $0.027 as of Jun 2026.$0$0.05$0.1$0.15$0.2Output on 11 Jun 2026: $0.201Out $0.201Input on 11 Jun 2026: $0.027In $0.027Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.027 $0.201 Imported from OpenRouter openrouter.ai

More from Meta

Report a problem