Context
1M
Max output
n/a
Serving providers
1
Cheapest input
$1.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial Analysis and Hugging Face leaderboardsSpeed & latency
Output speed
84 t/s
tokens / second
Time to first token
3.88s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| AIME 2026 | 97% | default | — | |
| CritPt | 5% | default | — | |
| EvalComputeProxy | 11767.0 | default | — | |
| GDPval | 1238.8 | default | — | |
| GPQA Diamond | 87% | default | — | |
| Humanity's Last Exam | 30% | default | — | |
| Long-Context Reasoning | 63% | default | — | |
| MMMU-Pro | 73% | default | — | |
| Omniscience | 2.0 | default | — | |
| SciCode | 46% | default | — | |
| SWE-bench Pro | 54% | default | — | |
| SWE-bench Verified | 78% | default | — | |
| TerminalBench v2.1 | 55% | default | — | |
| τ-bench Banking | 24% | default | — |
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| OpenRouterCheapest | $1.00 | $4.05 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Cost calculator
estimate your monthly spend on this model$524
estimated / month
Model IDs
copy the exact identifier for your platformthinkingmachines/inkling
Frequently asked questions
Inkling pricing, context and availabilityHow much does Inkling cost?
Inkling costs $1.00 per 1M input tokens and $4.05 per 1M output tokens at its cheapest provider via OpenRouter.
What is the context window of Inkling?
Inkling accepts up to 1M tokens of context.
Other Other models
compare pricing across the Other lineupMythoMax 13B$0.08/1M · 4 providersZephyr 7B Beta$0.15/1M · 2 providersKAT-Coder-Pro V2$0.30/1M · 2 providersMuse Spark 1.1$1.25/1M · 2 providersalia-40b-instruct_q8_0$0.00/1M · 1 providerApertus 70B Instruct$0.00/1M · 1 providerApertus 8B Instruct$0.00/1M · 1 providerBody Builder (beta)$0.00/1M · 1 provider
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI