Provider
MiniMax API pricing: every model compared
MiniMax offers 8 models through its API in our data, priced from $0.21 to $0.55 per million input tokens. The newest is MiniMax M3, first listed on 2026-05-31.
- Models
- 8
- Cheapest input
- $0.21
- Priciest input
- $0.55
- Largest context
- 1,000,192
- With caching
- 5 of 8
- Open weights
- 6 of 8
Cheapest
MiniMax M2.7MiniMax$0.21 in · $0.84 outLargest context
MiniMax-01MiniMax1,000,192 tokensNewest
MiniMax M3MiniMax2026-05-31Line-up
Every MiniMax model and price
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| MiniMax M3 | $0.30 | $1.20 | 524,288 | 2026-05-31 |
| MiniMax M2.7 | $0.21 | $0.84 | 196,608 | 2026-03-18 |
| MiniMax M2.5 | $0.27 | $1.08 | 200,000 | 2026-02-12 |
| MiniMax M2-her | $0.30 | $1.20 | 65,536 | 2026-01-23 |
| MiniMax M2.1 | $0.30 | $1.20 | 204,800 | 2025-12-23 |
| MiniMax M2 | $0.30 | $1.20 | 196,608 | 2025-10-23 |
| MiniMax M1 | $0.55 | $2.20 | 1,000,000 | 2025-06-17 |
| MiniMax-01 | $0.20 | $1.10 | 1,000,192 | 2025-01-15 |
Prices in US dollars per 1M tokens, newest first. Release dates are when our source first listed the model.
Overview
MiniMax’s pricing at a glance
The spread between MiniMax’s cheapest and most expensive model is large: MiniMax M1 costs 3× more per input token than MiniMax M2.7. Picking the smallest model that does the job well is usually the biggest saving available, ahead of any discount.
5 of 8 of the line-up has a cached-input price, which helps chatbots and agents that resend the same instructions. No batch prices are published in our data. 2 of 8 accept images as input.
FAQ
MiniMax API questions
How much does the MiniMax API cost?
MiniMax's 8 models range from $0.21 to $0.55 per million input tokens, and output costs more than input on almost every model. The exact bill depends on your token volumes, which you can estimate in the LLM cost calculator.
What is MiniMax's cheapest model?
By blended price (three parts input to one part output) it is MiniMax M2.7, at $0.21 input and $0.84 output per million tokens. The most expensive is MiniMax M1.
Which MiniMax model has the largest context window?
MiniMax-01, with 1,000,192 tokens per request.
Does MiniMax offer prompt caching or batch discounts?
In our data, 5 of 8 of MiniMax's models have a published cached-input price and none have a published batch price. Both can cut costs substantially for repeated prompts or work that can wait.