Context
n/a
Max output
n/a
Serving providers
0
Cheapest input
n/a /1M

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
104 t/s
tokens / second
Time to first token
0.33s
latency
Headline indices
36.5
of 100
3.1
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy117.7default
GDPval542.8default
GPQA Diamond76%default
Humanity's Last Exam11%default
IFBench58%default
Long-Context Reasoning36%default
Omniscience-48.6default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.2default
SciCode38%default
SWE-bench Pro40%default
SWE-bench Verified68%default
TerminalBench Hard31%default
TerminalBench v2.136%default
τ-bench Banking6%default
τ²-bench37%default

Pricing by serving provider

availability
No public price yet. Not offered on-demand at any tracked provider.

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Other Cohere models

compare pricing across the Cohere lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI