Llama Guard 3 8B
Meta · Released Feb 2025
An 8-billion-parameter content safety classifier from Meta that detects harmful content in text across multiple risk categories.
- Strengths
- Designed specifically for content moderation tasks, evaluating safety across categories like violence, illegal activity, harassment, and self-harm.
- Best for
- Screening user inputs or model outputs for harmful content in production applications that need safety filtering.
- Limitations
- A specialized safety classifier, not a general-purpose language model — it should not be used for tasks like text generation, question-answering, or reasoning; Llama Guard 4 12B (released April 2025) is a newer successor with expanded capabilities.
Input / 1M
$0.484
Output / 1M
$0.03
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.484 | $0.03 | — | Imported from OpenRouter | openrouter.ai |