prices synced 2026-09-11

R1 Distill Qwen 32B

DeepSeek · Released Jan 2025

Compare

R1 Distill Qwen 32B is a 32-billion-parameter reasoning model distilled from DeepSeek R1, designed to run efficiently on smaller hardware while preserving chain-of-thought reasoning capabilities.

Strengths
The model combines reasoning-focused inference with a smaller parameter count, making it faster and cheaper to run than larger reasoning models while still tackling logic-heavy problems.
Best for
Deployments where reasoning ability matters but latency and computational cost need to stay low, or where hardware constraints rule out larger reasoning models.
Limitations
As a distilled model, it trades some reasoning depth and general capability against its smaller size; it sits below R1 Distill Llama 70B in the reasoning-distill tier and is better suited to narrower, more constrained tasks than full-scale reasoning models.

Input / 1M

$0.29

Output / 1M

$0.29

Cached input / 1M

Context window

32K

Price history

Price per 1M tokens over timeOutput $0.29; Input $0.29 as of Jun 2026.$0$0.1$0.2$0.3Output on 11 Jun 2026: $0.29Out $0.29Input on 11 Jun 2026: $0.29In $0.29Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.29 $0.29 Imported from OpenRouter openrouter.ai

More from DeepSeek

Report a problem