LLMCost
thinkingmachines

Thinking Machines: Inkling API Pricing

Thinking Machines: Inkling costs $1.00 per 1M input tokens and $4.05 per 1M output tokens (cache reads: $0.170). Context window: 1.0M. Prices verified 2026-08-23 · source

Input / 1M tokens
$1.00
Output / 1M tokens
$4.05
Context window
1.0M
Max output

What would Thinking Machines: Inkling cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$24.20
Coding agent5,0008,000 / 2,000$80.50
RAG pipeline30,0004,000 / 500$180.75
Doc summarizer2,00020,000 / 1,000$48.10

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About Thinking Machines: Inkling

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

Compare Thinking Machines: Inkling with