Mistral Small 3
Mistral · Released Jan 2025
Mistral Small 3 is a 24-billion-parameter language model designed for low-latency inference across common AI tasks.
- Strengths
- It prioritizes fast response times while maintaining reasonable capability on standard tasks like summarization, classification, and question-answering.
- Best for
- Latency-sensitive applications where you need a balance between speed and capability, such as real-time API endpoints or high-throughput batch processing.
- Limitations
- At 24B parameters, it has less capacity than larger models and may struggle with complex reasoning, specialized domain knowledge, or tasks requiring extensive context understanding.
Input / 1M
$0.05
Output / 1M
$0.08
Cached input / 1M
—
Context window
32K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.05 | $0.08 | — | Imported from OpenRouter | openrouter.ai |