Context
256K
Max output
128K
Serving providers
2
Cheapest input
$0.30 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
100 t/s
tokens / second
Time to first token
1.27s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 0% | default | — | |
| GDPval | 905.7 | default | — | |
| GPQA Diamond | 85% | default | — | |
| Humanity's Last Exam | 16% | default | — | |
| IFBench | 67% | default | — | |
| Long-Context Reasoning | 70% | default | — | |
| Omniscience | -22.5 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.4 | default | — | |
| SciCode | 38% | default | — | |
| TerminalBench Hard | 49% | default | — | |
| TerminalBench v2.1 | 70% | default | — | |
| τ-bench Banking | 5% | default | — | |
| τ²-bench | 89% | default | — |
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| NovitaCheapest | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| OpenRouterCheapest | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$156
estimated / month
Model IDs
copy the exact identifier for your platformnovita/kwaipilot/kat-coder-prokwaipilot/kat-coder-pro-v2
Frequently asked questions
KAT-Coder-Pro V2 pricing, context and availabilityHow much does KAT-Coder-Pro V2 cost?
KAT-Coder-Pro V2 costs $0.30 per 1M input tokens and $1.20 per 1M output tokens at its cheapest provider via Novita.
What is the context window of KAT-Coder-Pro V2?
KAT-Coder-Pro V2 accepts up to 256K tokens of context and can return up to 128K output tokens.
Which providers serve KAT-Coder-Pro V2?
KAT-Coder-Pro V2 is available from 2 serving providers, each with its own pricing and model ID. Novita is currently the cheapest.
What can KAT-Coder-Pro V2 do?
KAT-Coder-Pro V2 supports tool use and structured output.
Other Other models
compare pricing across the Other lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI