Mistral Nemo
Mistral · Released Jul 2024
A 12B parameter model with a 128k token context window built by Mistral in collaboration with NVIDIA, designed to balance efficiency with capability across multiple languages.
- Strengths
- Compact size makes it suitable for deployment on resource-constrained hardware while maintaining support for long contexts and multiple languages including English, French, German, Spanish, Italian, Portuguese, Chinese, and Japanese.
- Best for
- Applications requiring a smaller model footprint but needing to handle extended documents or conversations in multiple languages.
- Limitations
- At 12B parameters, it will have reduced reasoning depth and may struggle with complex multi-step tasks compared to larger models.
Input / 1M
$0.019
Output / 1M
$0.03
Cached input / 1M
—
Context window
131K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in decreased 5.0% out decreased 0.0%
- 90d
- in decreased 5.0% out decreased 0.0%
- 1y
- in decreased 5.0% out decreased 0.0%
- Since launch
- in decreased 5.0% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 16 Jul 2026 | $0.019 | $0.03 | — | Imported from OpenRouter | openrouter.ai |
| 15 Jul 2026 | $0.02 | $0.04 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.02 | $0.03 | — | Imported from OpenRouter | openrouter.ai |