Qwen3.8 2.4T A95B
Alibaba · Released Aug 2026
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Alibaba with 95 billion active parameters out of 2.4 trillion total parameters.
- Strengths
- The sparse architecture activates only the necessary parameters per token, reducing compute requirements compared to dense models of equivalent capacity.
- Best for
- Workloads requiring substantial reasoning capacity where inference efficiency matters, particularly when deployed on infrastructure that supports sparse computation.
- Limitations
- Open-weight sparse models may require specialized hardware or software to realize efficiency gains; the model sits between Qwen3.6 35B A3B and Qwen3.8 Max in capability, lacking the multimodal features or optimizations of the flagship Max variant.
Input / 1M
$2.00
Output / 1M
$6.00
Cached input / 1M
$0.2
Context window
262K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 12 Aug 2026 | $2.00 | $6.00 | $0.2 | Imported from OpenRouter | openrouter.ai |