prices synced 2026-07-28

Llama 3 8B Instruct

Meta · Released Apr 2024

Compare

Meta's 8-billion-parameter instruction-tuned language model from the Llama 3 series, with an 8,192-token context window.

Strengths
Offers a balance of model capability and inference speed for general text understanding and generation tasks at a compact parameter count.
Best for
Low-latency applications and resource-constrained deployments where model size matters more than context length or reasoning depth.
Limitations
Smaller context window (8k tokens) and fewer parameters than later Llama 3.1 and 3.3 releases, making it less suitable for document-heavy tasks or complex reasoning; newer open-weight alternatives in Meta's lineup offer significantly better performance.

Input / 1M

$0.14

Output / 1M

$0.14

Cached input / 1M

Context window

8K

Price history

Price per 1M tokens over timeOutput $0.14; Input $0.14 as of Jun 2026.$0$0.05$0.1$0.15Output on 11 Jun 2026: $0.14Out $0.14Input on 11 Jun 2026: $0.14In $0.14Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.14 $0.14 Imported from OpenRouter openrouter.ai

More from Meta

Report a problem