Llama Guard 4 12B
Meta · Multimodal · Released Apr 2025
Llama Guard 4 12B is Meta's content safety classifier, built to detect and evaluate harmful content in both user prompts and model outputs.
- Strengths
- Operates at scale with a 163k-token context window and improved safety detection over its predecessor Llama Guard 3 8B.
- Best for
- Filtering harmful content in moderation pipelines and safety-critical applications where prompt and output evaluation is required.
- Limitations
- Designed as a safety classifier rather than a general-purpose language model; not intended for generation or reasoning tasks.
Input / 1M
$0.18
Output / 1M
$0.18
Cached input / 1M
—
Context window
163K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.18 | $0.18 | — | Imported from OpenRouter | openrouter.ai |