API pricing
Thinking Machines Lab API pricing
Thinking Machines Lab API prices run from $0.45 per million input tokens (Inkling-Small) to $1.87 (Inkling). Its highest-ranked model, Inkling-Small, costs $0.45 input and $1.20 output per million tokens.
Last verified
| Cached input | ||||||
|---|---|---|---|---|---|---|
| Inkling-Small (open weights) | $0.45 | $1.20 | $0.10 | $0.64 | 524K | 46.5 |
| Inkling (open weights) | $1.87 | $4.68 | $0.37 | $2.57 | 66K | 44.1 |
Prices on other platforms
The same models are often sold through clouds and resellers at different rates.
| Model | Route | Input | Output | Checked |
|---|---|---|---|---|
| Inkling-Small | deepinfra | $0.45 | $1.20 | 2026-10-10 |
| openrouter | $0.45 | $1.20 | 2026-10-10 | |
| Inkling | deepinfra | $0.95 | $4.05 | 2026-10-10 |
| fireworks | $1 | $4.05 | 2026-10-10 | |
| openrouter | $1 | $4.05 | 2026-10-10 | |
| thinking-machines | $1.87 | $4.68 | 2026-10-10 | |
| together | $1 | $4.05 | 2026-10-10 |
Frequently asked questions
How much does the Thinking Machines Lab API cost?
Thinking Machines Lab API prices run from $0.45 per million input tokens (Inkling-Small) to $1.87 (Inkling). Its highest-ranked model, Inkling-Small, costs $0.45 input and $1.20 output per million tokens.
Does Thinking Machines Lab discount cached input?
Yes. Cached input is billed at a lower rate on 2 of the 2 priced models; the cached rate is listed next to each model.