Context
131K
Max output
16K
Serving providers
2
Cheapest input
$0.55 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
40 t/s
tokens / second
Time to first token
1.29s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 0% | default | — | |
| EvalComputeProxy | 7.4 | default | — | |
| GPQA Diamond | 77% | default | — | |
| Humanity's Last Exam | 7% | default | — | |
| IFBench | 41% | default | — | |
| LiveCodeBench | 56% | default | — | |
| Long-Context Reasoning | 53% | default | — | |
| MMLU-Pro | 82% | default | — | |
| Omniscience | -28.3 | default | — | |
| Omniscience Accuracy | 0.3 | default | — | |
| Omniscience Non Hallucination | 0.2 | default | — | |
| SciCode | 34% | default | — | |
| TerminalBench Hard | 16% | default | — | |
| τ²-bench | 61% | default | — |
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expandInput pricing runs from $0.55 to $0.57 per 1M tokens across 2 providers, so the dearest route costs 4% more than the cheapest for the same model.
| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Vercel AI GatewayCheapest | $0.55 | $2.20 | 1 | ||||||||||||||||
| |||||||||||||||||||
| OpenRouter | $0.57 | $2.30 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$286
estimated / month
Model IDs
copy the exact identifier for your platformvercel_ai_gateway/moonshotai/kimi-k2moonshotai/kimi-k2
Frequently asked questions
Kimi K2 0711 pricing, context and availabilityHow much does Kimi K2 0711 cost?
Kimi K2 0711 costs $0.55 per 1M input tokens and $2.20 per 1M output tokens at its cheapest provider via Vercel AI Gateway. Across 2 serving providers, input prices range from $0.55 to $0.57 per 1M tokens.
What is the context window of Kimi K2 0711?
Kimi K2 0711 accepts up to 131K tokens of context and can return up to 16K output tokens.
Which providers serve Kimi K2 0711?
Kimi K2 0711 is available from 2 serving providers, each with its own pricing and model ID. Vercel AI Gateway is currently the cheapest.
What can Kimi K2 0711 do?
Kimi K2 0711 supports tool use.
Other Moonshot models
compare pricing across the Moonshot lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI