T
Compare
Inkling (batch)
Thinking Machines · Multimodal · Released Jul 2026
Open-weight multimodal mixture-of-experts model with 41B active parameters from 975B total and a 524K token context window.
- Strengths
- Handles multimodal inputs, supports long-context reasoning with its 524K token window, and uses mixture-of-experts routing to activate only a portion of its parameters per request.
- Best for
- General-purpose reasoning tasks, coding, agentic workflows, and tool-use systems that benefit from larger capacity than Inkling Small.
- Limitations
- As an open-weight model, it requires self-hosting or managed deployment; the larger parameter count and context window increase computational requirements compared to Inkling Small.
Input / 1M
$1.00
Output / 1M
$4.05
Cached input / 1M
$0.17
Context window
524K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in increased 100.0% out increased 100.0%
- 1y
- in increased 100.0% out increased 100.0%
- Since launch
- in increased 100.0% out increased 100.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 19 Aug 2026 | $1.00 | $4.05 | $0.17 | Imported from OpenRouter | openrouter.ai |
| 6 Aug 2026 | $0.5 | $2.02 | $0.085 | Imported from OpenRouter | openrouter.ai |