prices synced 2026-08-12

Qwen3.8 2.4T A95B

Alibaba · Released Aug 2026

Compare

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Alibaba with 95 billion active parameters out of 2.4 trillion total parameters.

Strengths
The sparse architecture activates only the necessary parameters per token, reducing compute requirements compared to dense models of equivalent capacity.
Best for
Workloads requiring substantial reasoning capacity where inference efficiency matters, particularly when deployed on infrastructure that supports sparse computation.
Limitations
Open-weight sparse models may require specialized hardware or software to realize efficiency gains; the model sits between Qwen3.6 35B A3B and Qwen3.8 Max in capability, lacking the multimodal features or optimizations of the flagship Max variant.

Input / 1M

$2.00

Output / 1M

$6.00

Cached input / 1M

$0.2

Context window

262K

Price history

Price per 1M tokens over timeOutput $6.00; Input $2.00 as of Aug 2026.$0$2.00$4.00$6.00Output on 12 Aug 2026: $6.00Out $6.00Input on 12 Aug 2026: $2.00In $2.00Aug 2026

Snapshots

Effective Input Output Cached in Note Source
12 Aug 2026 $2.00 $6.00 $0.2 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem