thinkingmachines
Thinking Machines: Inkling (batch) API Pricing
Thinking Machines: Inkling (batch) costs $1.00 per 1M input tokens and $4.05 per 1M output tokens (cache reads: $0.170). Context window: 524K. Prices verified 2026-08-23 · source
Input / 1M tokens
$1.00
Output / 1M tokens
$4.05
Context window
524K
Max output
—
What would Thinking Machines: Inkling (batch) cost you per month?
| Workload | Requests/mo | Avg tokens (in/out) | Est. monthly cost |
|---|---|---|---|
| Light chatbot | 10,000 | 800 / 400 | $24.20 |
| Coding agent | 5,000 | 8,000 / 2,000 | $80.50 |
| RAG pipeline | 30,000 | 4,000 / 500 | $180.75 |
| Doc summarizer | 2,000 | 20,000 / 1,000 | $48.10 |
Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.
About Thinking Machines: Inkling (batch)
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...