T
Compare
Inkling Small
Thinking Machines · Multimodal · Released Jul 2026
Open-weight multimodal mixture-of-experts model with 12B active parameters from 276B total and 524K token context.
- Strengths
- Efficient inference from a sparse parameter architecture while maintaining multimodal capabilities across long contexts.
- Best for
- Applications where reduced computational requirements and faster latency are prioritized over the scale of the full Inkling model.
- Limitations
- Smaller active parameter count compared to Inkling means reduced reasoning depth and knowledge capacity for complex tasks.
Input / 1M
$0.45
Output / 1M
$1.20
Cached input / 1M
$0.1
Context window
524K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in decreased 22.4% out decreased 16.7%
- 1y
- in decreased 22.4% out decreased 16.7%
- Since launch
- in decreased 22.4% out decreased 16.7%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 7 Aug 2026 | $0.45 | $1.20 | $0.1 | Imported from OpenRouter | openrouter.ai |
| 1 Aug 2026 | $0.5 | $1.20 | $0.1 | Imported from OpenRouter | openrouter.ai |
| 31 Jul 2026 | $0.58 | $1.44 | $0.116 | Imported from OpenRouter | openrouter.ai |