MiniMax-01 is a multimodal model combining text generation and image understanding with 456 billion parameters and a 1 million token context window.
- Strengths
- Handles both text and image inputs with a very large context window, enabling long-form reasoning and document processing.
- Best for
- Multimodal tasks that require processing long documents or extended conversations with both text and visual content.
- Limitations
- As a large model with sparse activation, it may have higher latency and memory requirements compared to smaller alternatives for simple tasks.
Input / 1M
$0.2
Output / 1M
$1.10
Cached input / 1M
—
Context window
1M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.2 | $1.10 | — | Imported from OpenRouter | openrouter.ai |