prices synced 2026-09-11
M

MiniMax M2-her

MiniMax · Released Jan 2026

Compare

MiniMax M2-her is a large language model with a 65K token context window, positioned as a lightweight tier within MiniMax's M2 generation.

Strengths
Offers fast inference and low latency suitable for real-time applications while maintaining competent performance across general language tasks.
Best for
Latency-sensitive deployments and scenarios where smaller model size and faster response time outweigh the need for extended context or maximum reasoning capacity.
Limitations
The 65K token context window is narrower than later models in the lineup such as M2.5 and M2.7, which support 204K tokens, limiting its ability to handle long documents or complex multi-turn conversations.

Input / 1M

$0.3

Output / 1M

$1.20

Cached input / 1M

$0.03

Context window

65K

Price history

Price per 1M tokens over timeOutput $1.20; Input $0.3 as of Jun 2026.$0$0.5$1.00Output on 11 Jun 2026: $1.20Out $1.20Input on 11 Jun 2026: $0.3In $0.3Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.3 $1.20 $0.03 Imported from OpenRouter openrouter.ai

More from MiniMax

Report a problem