Mistral Large 3 2512
retiredMistral · Multimodal · Released Dec 2025
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
- Strengths
- Balances reasoning capability with efficient inference through its sparse MoE architecture, activating only necessary parameters per token.
- Best for
- Complex reasoning tasks, long-context document processing, and applications where inference efficiency matters alongside capability.
- Limitations
- MoE models can have less predictable performance characteristics than dense models, and may underperform on tasks where all parameters contribute equally.
Input / 1M
$0.5
Output / 1M
$1.50
Cached input / 1M
$0.05
Context window
262K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.5 | $1.50 | $0.05 | Imported from OpenRouter | openrouter.ai |