R1 Distill Qwen 32B
DeepSeek · Released Jan 2025
R1 Distill Qwen 32B is a 32-billion-parameter reasoning model distilled from DeepSeek R1, designed to run efficiently on smaller hardware while preserving chain-of-thought reasoning capabilities.
- Strengths
- The model combines reasoning-focused inference with a smaller parameter count, making it faster and cheaper to run than larger reasoning models while still tackling logic-heavy problems.
- Best for
- Deployments where reasoning ability matters but latency and computational cost need to stay low, or where hardware constraints rule out larger reasoning models.
- Limitations
- As a distilled model, it trades some reasoning depth and general capability against its smaller size; it sits below R1 Distill Llama 70B in the reasoning-distill tier and is better suited to narrower, more constrained tasks than full-scale reasoning models.
Input / 1M
$0.29
Output / 1M
$0.29
Cached input / 1M
—
Context window
32K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.29 | $0.29 | — | Imported from OpenRouter | openrouter.ai |