Context
1M
Max output
524K
Serving providers
8
Cheapest input
$0.23 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
108 t/s
tokens / second
Time to first token
1.08s
latency
Headline indices
58.6
of 100
36.1
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
Aa Analyst Agent0.1default
AA-Briefcase31%
Automation Bench0.2default
CritPt4%default
EvalComputeProxy4190.3default
GDPval1386.9default
GPQA Diamond93%default
Harvey Lab0.9default
Humanity's Last Exam39%default
IFBench83%default
Long-Context Reasoning80%default
Mlcr Overall0.2default
MMMU-Pro79%default
Omniscience1.4default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.8default
SciCode45%default
SWE-bench Pro59%default
SWE-bench Verified81%default
TerminalBench Hard42%default
TerminalBench v2.165%default
τ-bench Banking15%default
τ²-bench89%default

Pricing detail

cost beyond the standard rate · source: models.dev
Long context (>200K)
$0.60 / $2.40
input / output · per 1M

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.23 to $0.30 per 1M tokens across 8 providers, so the dearest route costs 30% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
WandbCheapest$0.23$0.961
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.23$0.96$0.05wandb/MiniMaxAI/MiniMax-M3
DeepInfra$0.28$1.101
Fireworks AI$0.30$1.201
Minimax$0.30$1.201
Together AI$0.30$1.201
Tencent$0.30$1.201
Novita$0.30$1.201
OpenRouter$0.30$1.201

Price history

input + output $/1M since we started tracking
Input Output
Input down 62% since first tracked
$0.000$1.00$2.00$3.00Jun 5Jun 18Jul 19Aug 28$0.960$0.230

Cost calculator

estimate your monthly spend on this model
$123
estimated / month

Model IDs

copy the exact identifier for your platform
wandb/MiniMaxAI/MiniMax-M3deepinfra/MiniMaxAI/MiniMax-M3fireworks_ai/accounts/fireworks/models/minimax-m3minimax/MiniMax-M3together_ai/MiniMaxAI/MiniMax-M3tencent/minimax-m3novita/minimax/minimax-m3minimax/minimax-m3

Frequently asked questions

MiniMax M3 pricing, context and availability

How much does MiniMax M3 cost?

MiniMax M3 costs $0.23 per 1M input tokens and $0.96 per 1M output tokens at its cheapest provider via Wandb. Across 8 serving providers, input prices range from $0.23 to $0.30 per 1M tokens.

What is the context window of MiniMax M3?

MiniMax M3 accepts up to 1M tokens of context and can return up to 524K output tokens.

Which providers serve MiniMax M3?

MiniMax M3 is available from 8 serving providers, each with its own pricing and model ID. Wandb is currently the cheapest.

What can MiniMax M3 do?

MiniMax M3 supports image input (vision), tool use, reasoning, prompt caching and structured output.

Other MiniMax models

compare pricing across the MiniMax lineup

Popular comparisons

head-to-head pages featuring MiniMax M3
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI