T
Compare
Inkling Small
Thinking Machines · Multimodal · Released Jul 2026
Open-weight multimodal mixture-of-experts model with 12B active parameters from 276B total and 524K token context.
- Strengths
- Efficient inference from a sparse parameter architecture while maintaining multimodal capabilities across long contexts.
- Best for
- Applications where reduced computational requirements and faster latency are prioritized over the scale of the full Inkling model.
- Limitations
- Smaller active parameter count compared to Inkling means reduced reasoning depth and knowledge capacity for complex tasks.
Input / 1M
$0.58
Output / 1M
$1.44
Cached input / 1M
$0.116
Context window
524K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 31 Jul 2026 | $0.58 | $1.44 | $0.116 | Imported from OpenRouter | openrouter.ai |