Context
1M
Max output
n/a
Serving providers
1
Cheapest input
$3.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
62 t/s
tokens / second
Time to first token
8.97s
latency
Headline indices
76.2
of 100
50.1
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
APEX Agents41%default
CritPt23%default
EvalComputeProxy853190.9default
GDPval1683.7default
GPQA Diamond94%default
Humanity's Last Exam44%default
IT-Bench SRE48%default
Long-Context Reasoning75%default
MMMU-Pro81%default
Omniscience18.4default
SciCode59%default
TerminalBench v2.185%default
τ-bench Banking33%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$3.00$15.001
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$3.00$15.00$0.30moonshotai/kimi-k3

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$1,800
estimated / month

Model IDs

copy the exact identifier for your platform
moonshotai/kimi-k3

Frequently asked questions

Kimi K3 pricing, context and availability

How much does Kimi K3 cost?

Kimi K3 costs $3.00 per 1M input tokens and $15.00 per 1M output tokens at its cheapest provider via OpenRouter.

What is the context window of Kimi K3?

Kimi K3 accepts up to 1M tokens of context.

Other Moonshot models

compare pricing across the Moonshot lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI