Inkling API Pricing & Cost Calculator

Inkling is the wild-card workhorse for buyers wanting an alternative to OpenAI, Anthropic, Google, and the Chinese labs. Pricing is competitive for high-volume document work with a real 1M context.

Inkling is Thinking Machines Lab's flagship model and their first broadly available API release. It positions at workhorse pricing alongside GPT-5.6 Terra, with a 1-million-token context for agentic workloads.

By TechCompare · Updated

Input tokens
50,000
50% cached
Output tokens
3,000
per request
Volume
500 / monthly
Standard API

Calculator

Cost Comparison

Based on 50,000 input tokens (50% cached), 3,000 output tokens, and 500 requests.

How this is calculated

Inkling is priced at $1.00 per million input tokens and $4.05 per million output tokens, with an 83% prompt caching discount ($0.17 per million cached reads). That is half of GPT-5.6 Terra's input list rate ($2.00) while sitting just below GLM-5.2 ($4.40) and well below Terra on output.

Verdict

Inkling is the wild-card workhorse for buyers wanting an alternative to OpenAI, Anthropic, Google, and the Chinese labs. Pricing is competitive for high-volume document work with a real 1M context.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

Who is Thinking Machines Lab and why is Inkling notable?
Thinking Machines Lab is a recent AI startup, and Inkling is their first broadly available model on OpenRouter. It positions as a workhorse mid-tier alternative to GPT-5.6 Terra and GLM-5.2 with a real 1M context window.