prices synced 2026-09-21
T

Inkling (batch)

Thinking Machines · Multimodal · Released Jul 2026

Compare

Open-weight multimodal mixture-of-experts model with 41B active parameters from 975B total and a 524K token context window.

Strengths
Handles multimodal inputs, supports long-context reasoning with its 524K token window, and uses mixture-of-experts routing to activate only a portion of its parameters per request.
Best for
General-purpose reasoning tasks, coding, agentic workflows, and tool-use systems that benefit from larger capacity than Inkling Small.
Limitations
As an open-weight model, it requires self-hosting or managed deployment; the larger parameter count and context window increase computational requirements compared to Inkling Small.

Input / 1M

$1.00

Output / 1M

$4.05

Cached input / 1M

$0.17

Context window

524K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $2.02 to $4.05; Input went from $0.5 to $1.00 between Aug 2026 and Aug 2026.$0$1.00$2.00$3.00$4.00Output on 6 Aug 2026: $2.02Output on 19 Aug 2026: $4.05Out $4.05Input on 6 Aug 2026: $0.5Input on 19 Aug 2026: $1.00In $1.00Aug 2026Aug 2026

Price change

30d
in out
90d
in increased 100.0% out increased 100.0%
1y
in increased 100.0% out increased 100.0%
Since launch
in increased 100.0% out increased 100.0%

Snapshots

Effective Input Output Cached in Note Source
19 Aug 2026 $1.00 $4.05 $0.17 Imported from OpenRouter openrouter.ai
6 Aug 2026 $0.5 $2.02 $0.085 Imported from OpenRouter openrouter.ai

More from Thinking Machines

Report a problem