Context
1M
Max output
197K
Serving providers
6
Cheapest input
$0.27 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: ARC Prize, Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
100 t/s
tokens / second
Time to first token
1.65s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
ARC-AGI-25%default$0.170
CritPt1%default
EvalComputeProxy103.3default
GPQA Diamond85%default
Humanity's Last Exam21%default
IFBench72%default
Long-Context Reasoning72%default
Omniscience-38.9default
Omniscience Accuracy0.3default
Omniscience Non Hallucination0.1default
SciCode43%default
SWE-bench Pro55%default
SWE-bench Verified76%default
TerminalBench Hard35%default
τ²-bench95%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.27 to $0.30 per 1M tokens across 6 providers, so the dearest route costs 11% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.27$1.081
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.27$1.08$0.03minimax/minimax-m2.5
Baseten$0.30$1.201
Minimax$0.30$1.201
Wandb$0.30$1.201
Novita$0.30$1.201
Tensormesh$0.30$1.201

Price history

input + output $/1M since we started tracking
Input Output
Input down 10% since first tracked
$0.000$0.500$1.00$1.50Feb 12Jun 10Aug 6Aug 28$1.08$0.270

Cost calculator

estimate your monthly spend on this model
$140
estimated / month

Model IDs

copy the exact identifier for your platform
minimax/minimax-m2.5baseten/MiniMaxAI/MiniMax-M2.5minimax/MiniMax-M2.5wandb/MiniMaxAI/MiniMax-M2.5novita/minimax/minimax-m2.5tensormesh/MiniMaxAI/MiniMax-M2.5

Frequently asked questions

MiniMax M2.5 pricing, context and availability

How much does MiniMax M2.5 cost?

MiniMax M2.5 costs $0.27 per 1M input tokens and $1.08 per 1M output tokens at its cheapest provider via OpenRouter. Across 6 serving providers, input prices range from $0.27 to $0.30 per 1M tokens.

What is the context window of MiniMax M2.5?

MiniMax M2.5 accepts up to 1M tokens of context and can return up to 197K output tokens.

Which providers serve MiniMax M2.5?

MiniMax M2.5 is available from 6 serving providers, each with its own pricing and model ID. OpenRouter is currently the cheapest.

What can MiniMax M2.5 do?

MiniMax M2.5 supports tool use, reasoning, prompt caching and structured output.

Other MiniMax models

compare pricing across the MiniMax lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI