Llama 3.3 70B Instruct
Meta · Released Dec 2024
A 70-billion-parameter instruction-tuned language model from Meta with a 131,072-token context window, released in December 2024.
- Strengths
- Handles long-context tasks and complex reasoning across diverse domains with improved performance over Llama 3.1 70B Instruct.
- Best for
- General-purpose text generation, summarization of lengthy documents, and multi-step reasoning tasks that benefit from a large model and extended context.
- Limitations
- Requires more compute and latency than smaller models like Llama 3.2 3B or 11B, making it less suitable for latency-critical or resource-constrained applications.
Input / 1M
$0.1
Output / 1M
$0.32
Cached input / 1M
—
Context window
131K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in decreased 0.0% out decreased 0.0%
- 90d
- in decreased 0.0% out decreased 0.0%
- 1y
- in decreased 0.0% out decreased 0.0%
- Since launch
- in decreased 0.0% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 3 Sep 2026 | $0.1 | $0.32 | — | Imported from OpenRouter | openrouter.ai |
| 26 Aug 2026 | $0.71 | $0.71 | $0.71 | Imported from OpenRouter | openrouter.ai |
| 4 Aug 2026 | $0.1 | $0.32 | — | Imported from OpenRouter | openrouter.ai |
| 15 Jul 2026 | $0.13 | $0.4 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.1 | $0.32 | — | Imported from OpenRouter | openrouter.ai |