Context
1M
Max output
524K
Serving providers
8
Cheapest input
$0.23 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial Analysis and Hugging Face leaderboardsSpeed & latency
Output speed
108 t/s
tokens / second
Time to first token
1.08s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| Aa Analyst Agent | 0.1 | default | — | |
| AA-Briefcase | 31% | — | — | |
| Automation Bench | 0.2 | default | — | |
| CritPt | 4% | default | — | |
| EvalComputeProxy | 4190.3 | default | — | |
| GDPval | 1386.9 | default | — | |
| GPQA Diamond | 93% | default | — | |
| Harvey Lab | 0.9 | default | — | |
| Humanity's Last Exam | 39% | default | — | |
| IFBench | 83% | default | — | |
| Long-Context Reasoning | 80% | default | — | |
| Mlcr Overall | 0.2 | default | — | |
| MMMU-Pro | 79% | default | — | |
| Omniscience | 1.4 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.8 | default | — | |
| SciCode | 45% | default | — | |
| SWE-bench Pro | 59% | default | — | |
| SWE-bench Verified | 81% | default | — | |
| TerminalBench Hard | 42% | default | — | |
| TerminalBench v2.1 | 65% | default | — | |
| τ-bench Banking | 15% | default | — | |
| τ²-bench | 89% | default | — |
Pricing detail
cost beyond the standard rate · source: models.devLong context (>200K)
$0.60 / $2.40
input / output · per 1M
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expandInput pricing runs from $0.23 to $0.30 per 1M tokens across 8 providers, so the dearest route costs 30% more than the cheapest for the same model.
| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| WandbCheapest | $0.23 | $0.96 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| DeepInfra | $0.28 | $1.10 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Fireworks AI | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Minimax | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Together AI | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Tencent | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Novita | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| OpenRouter | $0.30 | $1.20 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Input down 62% since first tracked
Cost calculator
estimate your monthly spend on this model$123
estimated / month
Model IDs
copy the exact identifier for your platformwandb/MiniMaxAI/MiniMax-M3deepinfra/MiniMaxAI/MiniMax-M3fireworks_ai/accounts/fireworks/models/minimax-m3minimax/MiniMax-M3together_ai/MiniMaxAI/MiniMax-M3tencent/minimax-m3novita/minimax/minimax-m3minimax/minimax-m3
Frequently asked questions
MiniMax M3 pricing, context and availabilityHow much does MiniMax M3 cost?
MiniMax M3 costs $0.23 per 1M input tokens and $0.96 per 1M output tokens at its cheapest provider via Wandb. Across 8 serving providers, input prices range from $0.23 to $0.30 per 1M tokens.
What is the context window of MiniMax M3?
MiniMax M3 accepts up to 1M tokens of context and can return up to 524K output tokens.
Which providers serve MiniMax M3?
MiniMax M3 is available from 8 serving providers, each with its own pricing and model ID. Wandb is currently the cheapest.
What can MiniMax M3 do?
MiniMax M3 supports image input (vision), tool use, reasoning, prompt caching and structured output.
Other MiniMax models
compare pricing across the MiniMax lineupPopular comparisons
head-to-head pages featuring MiniMax M3 Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI