Context
n/a
Max output
n/a
Serving providers
0
Cheapest input
n/a /1M
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
64 t/s
tokens / second
Time to first token
2.57s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 4% | default | — | |
| EvalComputeProxy | 266.1 | default | — | |
| GDPval | 1150.4 | default | — | |
| GPQA Diamond | 85% | default | — | |
| Humanity's Last Exam | 27% | default | — | |
| IFBench | 67% | default | — | |
| Long-Context Reasoning | 68% | default | — | |
| MMMU-Pro | 75% | default | — | |
| Omniscience | -9.8 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.7 | default | — | |
| SciCode | 43% | default | — | |
| TerminalBench Hard | 42% | default | — | |
| TerminalBench v2.1 | 64% | default | — | |
| τ-bench Banking | 9% | default | — | |
| τ²-bench | 91% | default | — |
Pricing by serving provider
availabilityNo public price yet. Not offered on-demand at any tracked provider.
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Other Xiaomi models
compare pricing across the Xiaomi lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI