A 405-billion parameter reasoning model based on Llama 3.1 that can choose between direct answers and extended internal deliberation modes.
- Strengths
- Excels at complex reasoning tasks where it can allocate compute by deciding when extended deliberation helps versus when direct responses suffice.
- Best for
- Technical problem-solving, mathematical reasoning, and multi-step logic chains where the model can benefit from flexible reasoning depth.
- Limitations
- The hybrid reasoning approach requires careful prompt engineering to work effectively, and deliberation mode adds latency on tasks that don't need it.
Input / 1M
$1.00
Output / 1M
$3.00
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $1.00 | $3.00 | — | Imported from OpenRouter | openrouter.ai |