A 4-billion-parameter multimodal guardrail model from NVIDIA designed to moderate both inputs to and outputs from language and vision models.
- Strengths
- Efficient at detecting unsafe content across text and image modalities without blocking legitimate requests.
- Best for
- Filtering harmful content in production LLM and VLM deployments where model size and latency constraints require a lightweight safety classifier.
- Limitations
- At 4B parameters, it has lower capacity than NVIDIA's larger general-purpose models and may miss edge cases in complex safety scenarios.
Input / 1M
$0.2
Output / 1M
$0.2
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 4 Sep 2026 | $0.2 | $0.2 | — | Imported from OpenRouter | openrouter.ai |