Qwen2.5 7B Instruct
Alibaba · Released Oct 2024
A 7-billion-parameter instruction-tuned dense language model from Alibaba with a 32k-token context window.
- Strengths
- Runs efficiently on resource-constrained hardware while maintaining reasonable quality for general language tasks.
- Best for
- Latency-sensitive applications, edge deployments, and scenarios where model size and inference speed matter more than maximum capability.
- Limitations
- Significantly smaller than newer Alibaba releases like Qwen3 8B and larger variants, with reduced performance on complex reasoning and specialized tasks like code generation.
Input / 1M
$0.1
Output / 1M
$0.2
Cached input / 1M
—
Context window
32K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in increased 150.0% out increased 100.0%
- 1y
- in increased 150.0% out increased 100.0%
- Since launch
- in increased 150.0% out increased 100.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 30 Jul 2026 | $0.1 | $0.2 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.04 | $0.1 | — | Imported from OpenRouter | openrouter.ai |