Llama 3 8B Instruct
Meta · Released Apr 2024
Meta's 8-billion-parameter instruction-tuned language model from the Llama 3 series, with an 8,192-token context window.
- Strengths
- Offers a balance of model capability and inference speed for general text understanding and generation tasks at a compact parameter count.
- Best for
- Low-latency applications and resource-constrained deployments where model size matters more than context length or reasoning depth.
- Limitations
- Smaller context window (8k tokens) and fewer parameters than later Llama 3.1 and 3.3 releases, making it less suitable for document-heavy tasks or complex reasoning; newer open-weight alternatives in Meta's lineup offer significantly better performance.
Input / 1M
$0.14
Output / 1M
$0.14
Cached input / 1M
—
Context window
8K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.14 | $0.14 | — | Imported from OpenRouter | openrouter.ai |