MiniMax Pricing
MiniMax charges between $1.02 and $2.40 per 1 million output tokens, depending on the model. MiniMax M2 is the cheapest at $1.02/1M output and $0.255/1M input; MiniMax-M2.7-highspeed is the most expensive at $2.40/1M output. The largest context window is 1.0M tokens (MiniMax-01). Prices are USD, current as of July 20, 2026.
Last updated · synced weekly from the upstream model catalog
Cheapest input
$0.200/1M
MiniMax-01
Cheapest output
$1.02/1M
MiniMax M2
Longest context
1.0M
MiniMax-01
Models priced
11
Avg $1.70/1M output
Pricing by model
All prices in USD per 1 million tokens. All 11 priced models, cheapest first.
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
| MiniMax M2 | $0.255 | $1.02 | 205K |
| MiniMax-01 | $0.200 | $1.10 | 1.0M |
| MiniMax-M2.1 ReasoningTools | $0.300 | $1.20 | 205K |
| MiniMax-M2.5 ReasoningToolsCache | $0.300 | $1.20 | 205K |
| MiniMax-M2.7 ReasoningToolsCache | $0.300 | $1.20 | 205K |
| MiniMax-M3 VisionReasoningToolsCache | $0.300 | $1.20 | 1.0M |
| MiniMax M1 | $0.400 | $2.20 | 1.0M |
| Minimax M2.1 Lightning | $0.300 | $2.40 | 1.0M |
| Minimax M2.5 Lightning | $0.300 | $2.40 | 1.0M |
| MiniMax-M2.5-highspeed ReasoningToolsCache | $0.600 | $2.40 | 205K |
| MiniMax-M2.7-highspeed ReasoningToolsCache | $0.600 | $2.40 | 205K |
Frequently asked questions
How much does MiniMax cost per 1M tokens?
MiniMax pricing ranges from $1.02 to $2.40 per 1 million output tokens across 11 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.
What is the cheapest MiniMax model?
MiniMax M2 is the cheapest MiniMax model on output tokens at $1.02 per 1M, while MiniMax-01 is cheapest on input at $0.200 per 1M. Which wins for you depends on your input-to-output ratio.
What is the largest MiniMax context window?
MiniMax-01 has the largest context window in the MiniMax catalog at 1.0M input tokens, with up to 16K output tokens per response.
Does MiniMax support prompt caching?
Yes. MiniMax-M2.5 reads cached input at $0.030 per 1M tokens, against $0.300 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.
Which MiniMax models support reasoning or vision?
The MiniMax catalog includes 6 reasoning models, 1 vision models, 6 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.
How do I calculate my actual MiniMax bill?
Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.
Other providers
What will this actually cost you?
Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.
Open the free cost calculator