Context
n/a
Max output
n/a
Serving providers
0
Cheapest input
n/a /1M

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
76 t/s
tokens / second
Time to first token
2.87s
latency
Headline indices
73.1
of 100
56.4
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt11%default
GDPval (normalized)62%
GPQA Diamond92%default
Humanity's Last Exam38%default
Long-Context Reasoning77%default
MMMU-Pro80%default
Omniscience-9.7default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.5default
SciCode47%default
SWE-bench Pro63%default
TerminalBench v2.186%default
τ-bench Banking45%default

Pricing by serving provider

availability
No public price yet. Not offered on-demand at any tracked provider.

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI