Ministral 8B
Mistral · Released Oct 2024
Ministral 8B is an 8-billion-parameter model from Mistral with interleaved sliding-window attention for efficient inference and a 128k token context window.
- Strengths
- The sliding-window attention pattern reduces memory consumption and inference latency while maintaining a large context window.
- Best for
- Edge deployments and resource-constrained environments where inference speed and memory efficiency are priorities.
- Limitations
- At 8B parameters, it trades off reasoning capability and knowledge breadth compared to larger models in Mistral's lineup like Mistral Small 3 (24B) or Mistral Large 3 (41B active).
Input / 1M
$0.11
Output / 1M
$0.11
Cached input / 1M
—
Context window
128K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 20 Aug 2026 | $0.11 | $0.11 | — | Imported from OpenRouter | openrouter.ai |