prices synced 2026-07-28

Mistral Nemo

Mistral · Released Jul 2024

Compare

A 12B parameter model with a 128k token context window built by Mistral in collaboration with NVIDIA, designed to balance efficiency with capability across multiple languages.

Strengths
Compact size makes it suitable for deployment on resource-constrained hardware while maintaining support for long contexts and multiple languages including English, French, German, Spanish, Italian, Portuguese, Chinese, and Japanese.
Best for
Applications requiring a smaller model footprint but needing to handle extended documents or conversations in multiple languages.
Limitations
At 12B parameters, it will have reduced reasoning depth and may struggle with complex multi-step tasks compared to larger models.

Input / 1M

$0.019

Output / 1M

$0.03

Cached input / 1M

Context window

131K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $0.03 to $0.03; Input went from $0.02 to $0.019 between Jun 2026 and Jul 2026.$0$0.01$0.02$0.03$0.04Output on 11 Jun 2026: $0.03Output on 15 Jul 2026: $0.04Output on 16 Jul 2026: $0.03Out $0.03Input on 11 Jun 2026: $0.02Input on 15 Jul 2026: $0.02Input on 16 Jul 2026: $0.019In $0.019Jun 2026Jul 2026

Price change

30d
in decreased 5.0% out decreased 0.0%
90d
in decreased 5.0% out decreased 0.0%
1y
in decreased 5.0% out decreased 0.0%
Since launch
in decreased 5.0% out decreased 0.0%

Snapshots

Effective Input Output Cached in Note Source
16 Jul 2026 $0.019 $0.03 Imported from OpenRouter openrouter.ai
15 Jul 2026 $0.02 $0.04 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.02 $0.03 Imported from OpenRouter openrouter.ai

More from Mistral

Report a problem