prices synced 2026-09-11

Qwen2.5 72B Instruct

Alibaba · Released Sep 2024

Compare

Qwen2.5 72B Instruct is a 72-billion-parameter instruction-tuned dense language model from Alibaba with a 32,768-token context window.

Strengths
It handles complex reasoning, multi-turn conversations, and nuanced instruction-following tasks at a scale larger than smaller Qwen2.5 models.
Best for
Production workloads that require strong instruction-following and reasoning where the smaller Qwen2.5 7B model lacks capacity.
Limitations
Newer Qwen3 mixture-of-experts models offer better parameter efficiency, and it is purely text-based unlike Qwen2.5 VL 72B Instruct which handles multimodal input.

Input / 1M

$0.36

Output / 1M

$0.4

Cached input / 1M

Context window

32K

Price history

Price per 1M tokens over timeOutput $0.4; Input $0.36 as of Jun 2026.$0$0.1$0.2$0.3$0.4Output on 11 Jun 2026: $0.4Out $0.4Input on 11 Jun 2026: $0.36In $0.36Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.36 $0.4 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem