R1 Distill Qwen 32B
DeepSeek · Released Jan 2025
A 32B parameter model distilled from DeepSeek R1 using Qwen 2.5 as the base architecture, trained to replicate reasoning capabilities of the larger R1 model.
- Strengths
- Delivers strong performance on reasoning and problem-solving tasks while maintaining a manageable model size for inference.
- Best for
- Applications requiring multi-step reasoning, mathematical problem-solving, and code generation without the computational overhead of larger models.
- Limitations
- As a distilled model, it may not match the performance of the full DeepSeek R1 on highly complex reasoning tasks requiring extended chain-of-thought computation.
Input / 1M
$0.29
Output / 1M
$0.29
Cached input / 1M
—
Context window
32K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.29 | $0.29 | — | Imported from OpenRouter | openrouter.ai |