prices synced 2026-07-28

Llama 3.2 1B Instruct

Meta · Released Sep 2024

Compare

A 1-billion-parameter instruction-tuned language model from Meta, released in September 2024 as part of the Llama 3.2 family.

Strengths
Extremely lightweight and fast to run, making it suitable for latency-sensitive applications and edge deployment where computational resources are constrained.
Best for
Low-latency inference, mobile and edge devices, and applications where speed and minimal resource footprint take priority over reasoning depth.
Limitations
At 1B parameters, it has limited reasoning capability and context understanding compared to larger models in the Llama lineup; the 3B and 8B variants offer significantly better performance for complex tasks.

Input / 1M

$0.027

Output / 1M

$0.201

Cached input / 1M

Context window

60K

Price history

Price per 1M tokens over timeOutput $0.201; Input $0.027 as of Jun 2026.$0$0.05$0.1$0.15$0.2Output on 11 Jun 2026: $0.201Out $0.201Input on 11 Jun 2026: $0.027In $0.027Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.027 $0.201 Imported from OpenRouter openrouter.ai

More from Meta

Report a problem