Context
n/a
Max output
n/a
Serving providers
0
Cheapest input
n/a /1M

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
99 t/s
tokens / second
Time to first token
1.07s
latency
Headline indices
34.0
of 100
7.0
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%low
GDPval681.6medium
GPQA Diamond78%medium
Humanity's Last Exam9%medium
IFBench80%low
LiveCodeBench73%medium
Long-Context Reasoning63%low
Mlcr Overall0.0medium
MMLU-Pro83%medium
MMMU-Pro65%medium
Omniscience-40.5low
Omniscience Accuracy0.2medium
Omniscience Non Hallucination0.3low
SciCode43%medium
TerminalBench Hard24%medium
TerminalBench v2.130%medium
τ-bench Banking9%low
τ²-bench93%medium

Pricing by serving provider

availability
No public price yet. Not offered on-demand at any tracked provider.

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Other Amazon models

compare pricing across the Amazon lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI