prices synced 2026-09-11

Qwen3 32B

Alibaba · Released Apr 2025

Compare

A 32-billion-parameter dense language model from Alibaba with a 40,960-token context window.

Strengths
Offers a mid-range parameter count that balances capability and inference speed compared to smaller 8B and 14B models.
Best for
General-purpose tasks where context window constraints are not limiting and you want more capacity than the 14B tier.
Limitations
Falls between the denser Qwen3 14B and the mixture-of-experts efficiency of Qwen3 30B A3B, making it less clearly differentiated within the Alibaba lineup; smaller models like 8B or 14B may be more efficient for similar workloads.

Input / 1M

$0.08

Output / 1M

$0.28

Cached input / 1M

Context window

40K

Price history

Price per 1M tokens over timeOutput $0.28; Input $0.08 as of Jun 2026.$0$0.1$0.2$0.3Output on 11 Jun 2026: $0.28Out $0.28Input on 11 Jun 2026: $0.08In $0.08Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.08 $0.28 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem