Context
1M
Max output
131K
Serving providers
3
Cheapest input
$0.43 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial Analysis and Hugging Face leaderboardsSpeed & latency
Output speed
36 t/s
tokens / second
Time to first token
4.36s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Aa Analyst Agent | 0.2 | default | — | |||||||||||||
| AA-Briefcase | 19% | — | — | |||||||||||||
| APEX Agents | 2% | default | — | |||||||||||||
| Automation Bench | 0.2 | default | — | |||||||||||||
| CritPt | 4% | default | — | |||||||||||||
| ||||||||||||||||
| EvalComputeProxy | 1239.8 | default | — | |||||||||||||
| ||||||||||||||||
| GDPval | 1265.8 | default | — | |||||||||||||
| GPQA Diamond | 87% | default | — | |||||||||||||
| ||||||||||||||||
| Harvey Lab | 0.7 | default | — | |||||||||||||
| Humanity's Last Exam | 36% | default | — | |||||||||||||
| ||||||||||||||||
| IFBench | 80% | default | — | |||||||||||||
| ||||||||||||||||
| IT-Bench SRE | 38% | default | — | |||||||||||||
| Long-Context Reasoning | 78% | default | — | |||||||||||||
| ||||||||||||||||
| Mlcr Overall | 0.1 | default | — | |||||||||||||
| Omniscience | 3.3 | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience Accuracy | 0.3 | reasoning: false | — | |||||||||||||
| ||||||||||||||||
| Omniscience Non Hallucination | 0.8 | default | — | |||||||||||||
| ||||||||||||||||
| SciCode | 50% | default | — | |||||||||||||
| ||||||||||||||||
| SWE-bench Pro | 57% | default | — | |||||||||||||
| SWE-bench Verified | 79% | default | — | |||||||||||||
| TerminalBench Hard | 43% | default | — | |||||||||||||
| ||||||||||||||||
| TerminalBench v2.1 | 65% | default | — | |||||||||||||
| τ-bench Banking | 10% | default | — | |||||||||||||
| τ²-bench | 94% | default | — | |||||||||||||
| ||||||||||||||||
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expandInput pricing runs from $0.43 to $1.00 per 1M tokens across 3 providers, so the dearest route costs 130% more than the cheapest for the same model.
| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| OpenRouterCheapest | $0.43 | $0.87 | 1 | |||||||||||||
| ||||||||||||||||
| Novita | $0.52 | $1.04 | 1 | |||||||||||||
| ||||||||||||||||
| DeepInfra | $1.00 | $3.00 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Input down 56% since first tracked
Cost calculator
estimate your monthly spend on this model$157
estimated / month
Model IDs
copy the exact identifier for your platformxiaomi/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-prodeepinfra/XiaomiMiMo/MiMo-V2.5-Pro
Frequently asked questions
MiMo-V2.5-Pro pricing, context and availabilityHow much does MiMo-V2.5-Pro cost?
MiMo-V2.5-Pro costs $0.43 per 1M input tokens and $0.87 per 1M output tokens at its cheapest provider via OpenRouter. Across 3 serving providers, input prices range from $0.43 to $1.00 per 1M tokens.
What is the context window of MiMo-V2.5-Pro?
MiMo-V2.5-Pro accepts up to 1M tokens of context and can return up to 131K output tokens.
Which providers serve MiMo-V2.5-Pro?
MiMo-V2.5-Pro is available from 3 serving providers, each with its own pricing and model ID. OpenRouter is currently the cheapest.
What can MiMo-V2.5-Pro do?
MiMo-V2.5-Pro supports tool use, reasoning, prompt caching and structured output.
Other Xiaomi models
compare pricing across the Xiaomi lineupPopular comparisons
head-to-head pages featuring MiMo-V2.5-Pro Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI