prices synced 2026-07-28
X

MiMo-V2-Flash

Xiaomi · Released Dec 2025

Compare

MiMo-V2-Flash is a Mixture-of-Experts language model from Xiaomi with 309B total parameters and 15B active parameters, supporting a 262K token context window.

Strengths
The sparse MoE architecture keeps computational requirements low while leveraging a large parameter pool, and the extended context window enables processing of long documents in a single pass.
Best for
Applications requiring long-context processing and cost-efficient inference where sparse activation patterns are acceptable.
Limitations
As an open-source foundation model, it may require additional fine-tuning for specialized tasks, and sparse MoE models can have less predictable performance across all workload types compared to dense alternatives.

Input / 1M

$0.1

Output / 1M

$0.3

Cached input / 1M

$0.01

Context window

262K

Price history

Price per 1M tokens over timeOutput $0.3; Input $0.1 as of Jun 2026.$0$0.1$0.2$0.3Output on 11 Jun 2026: $0.3Out $0.3Input on 11 Jun 2026: $0.1In $0.1Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.1 $0.3 $0.01 Imported from OpenRouter openrouter.ai

More from Xiaomi

Report a problem