prices synced 2026-07-28

Mistral Small 3

Mistral · Released Jan 2025

Compare

Mistral Small 3 is a 24-billion-parameter language model designed for low-latency inference across common AI tasks.

Strengths
It prioritizes fast response times while maintaining reasonable capability on standard tasks like summarization, classification, and question-answering.
Best for
Latency-sensitive applications where you need a balance between speed and capability, such as real-time API endpoints or high-throughput batch processing.
Limitations
At 24B parameters, it has less capacity than larger models and may struggle with complex reasoning, specialized domain knowledge, or tasks requiring extensive context understanding.

Input / 1M

$0.05

Output / 1M

$0.08

Cached input / 1M

Context window

32K

Price history

Price per 1M tokens over timeOutput $0.08; Input $0.05 as of Jun 2026.$0$0.02$0.04$0.06$0.08Output on 11 Jun 2026: $0.08Out $0.08Input on 11 Jun 2026: $0.05In $0.05Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.05 $0.08 Imported from OpenRouter openrouter.ai

More from Mistral

Report a problem