prices synced 2026-07-28

Mistral Large 3 2512

retired

Mistral · Multimodal · Released Dec 2025

Compare

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Strengths
Balances reasoning capability with efficient inference through its sparse MoE architecture, activating only necessary parameters per token.
Best for
Complex reasoning tasks, long-context document processing, and applications where inference efficiency matters alongside capability.
Limitations
MoE models can have less predictable performance characteristics than dense models, and may underperform on tasks where all parameters contribute equally.

Input / 1M

$0.5

Output / 1M

$1.50

Cached input / 1M

$0.05

Context window

262K

Price history

Price per 1M tokens over timeOutput $1.50; Input $0.5 as of Jun 2026.$0$0.5$1.00$1.50Output on 11 Jun 2026: $1.50Out $1.50Input on 11 Jun 2026: $0.5In $0.5Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.5 $1.50 $0.05 Imported from OpenRouter openrouter.ai

More from Mistral

Report a problem