Llama Guard 4 12B
Meta · Multimodal · Released Apr 2025
A 12-billion-parameter content safety classifier from Meta that detects harmful content in text across multiple risk categories.
- Strengths
- Designed specifically for safety classification tasks, with a larger parameter count than its predecessor Llama Guard 3 8B to handle more nuanced risk detection.
- Best for
- Filtering user inputs and model outputs for harmful content in production systems that need reliable content moderation at scale.
- Limitations
- A specialized safety classifier, not a general-purpose language model; it is built to categorize risk rather than generate text or perform general reasoning tasks.
Input / 1M
$0.18
Output / 1M
$0.18
Cached input / 1M
—
Context window
163K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.18 | $0.18 | — | Imported from OpenRouter | openrouter.ai |