Llama 3.1 8B Instruct
Meta · Released Jul 2024
Meta's 8-billion-parameter instruction-tuned language model from the Llama 3.1 family, with a 131k token context window.
- Strengths
- Offers a good balance of capability and speed for a compact model, with support for long context windows that enable processing of substantial documents or conversation histories.
- Best for
- Tasks where response latency matters and you don't need the reasoning depth of larger models, such as conversational agents, customer support, and simple content generation.
- Limitations
- Smaller parameter count limits performance on complex reasoning, code generation, and tasks requiring broad knowledge compared to larger models like Llama 3.1 70B Instruct; a newer release like Llama 3.3 70B Instruct provides more capability in the instruction-tuned family.
Input / 1M
$0.05
Output / 1M
$0.08
Cached input / 1M
$0.025
Context window
131K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in increased 150.0% out increased 166.7%
- 90d
- in increased 150.0% out increased 166.7%
- 1y
- in increased 150.0% out increased 166.7%
- Since launch
- in increased 150.0% out increased 166.7%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 15 Jul 2026 | $0.05 | $0.08 | $0.025 | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.02 | $0.03 | — | Imported from OpenRouter | openrouter.ai |