prices synced 2026-09-11

Qwen2.5 7B Instruct

Alibaba · Released Oct 2024

Compare

A 7-billion-parameter instruction-tuned dense language model from Alibaba with a 32k-token context window.

Strengths
Runs efficiently on resource-constrained hardware while maintaining reasonable quality for general language tasks.
Best for
Latency-sensitive applications, edge deployments, and scenarios where model size and inference speed matter more than maximum capability.
Limitations
Significantly smaller than newer Alibaba releases like Qwen3 8B and larger variants, with reduced performance on complex reasoning and specialized tasks like code generation.

Input / 1M

$0.1

Output / 1M

$0.2

Cached input / 1M

Context window

32K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $0.1 to $0.2; Input went from $0.04 to $0.1 between Jun 2026 and Jul 2026.$0$0.05$0.1$0.15$0.2Output on 11 Jun 2026: $0.1Output on 30 Jul 2026: $0.2Out $0.2Input on 11 Jun 2026: $0.04Input on 30 Jul 2026: $0.1In $0.1Jun 2026Jul 2026

Price change

30d
in out
90d
in increased 150.0% out increased 100.0%
1y
in increased 150.0% out increased 100.0%
Since launch
in increased 150.0% out increased 100.0%

Snapshots

Effective Input Output Cached in Note Source
30 Jul 2026 $0.1 $0.2 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.04 $0.1 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem