MiniMax API pricing: every model, every rate
MiniMax ships a compact family with a notable gap between standard and highspeed variants — the latter bills several times more for lower latency.
MiniMax-M2. Rates checked 2026-09-16.Full rate card
6 billable models in total, 6 of them text models. On those text models: 0 carry a rate verified against the vendor's own published pricing, the rest are gateway rates only. USD per 1M tokens, checked 2026-09-16.
| Model | Input | Output | Out/In | Official (in / out) |
|---|---|---|---|---|
MiniMax-M2.7-highspeed | $2.1 | $8.4 | 4.0× | not verified |
MiniMax-M2 | $1.05 | $4.2 | 4.0× | not verified |
MiniMax-M2.1 | $1.05 | $4.2 | 4.0× | not verified |
MiniMax-M2.5 | $0.15 | $0.6 | 4.0× | not verified |
MiniMax-M2.7 | $0.15 | $0.6 | 4.0× | not verified |
MiniMax-M3 | $0.15 | $0.6 | 4.0× | not verified |
Output is billed at a multiple of input on most models. The Out/In column is that multiple — it matters more than the input price once output dominates your bill.
Across 6 priced text models the blended rate spans $0.75 (MiniMax-M2.5) to $10.5 (MiniMax-M2.7-highspeed) per 1M tokens — a 14× spread. Picking the wrong tier is usually the single most expensive mistake here.
What MiniMax costs at three usage levels
Priced on MiniMax-M2.5 at $0.15 in / $0.6 out; MiniMax-M2 at $1.05 in / $4.2 out per 1M tokens, with a 5:1 input-to-output ratio — roughly what an
interactive workload produces.
| Usage level | Tokens per month (in / out) | Lowest-cost tierMiniMax-M2.5 | Median-rate modelMiniMax-M2 |
|---|---|---|---|
| Light pilot or side project | 5M / 1M | $1.35 | $9.45 |
| Working one product in production | 50M / 10M | $13.50 | $94.50 |
| Heavy high-volume pipeline | 500M / 100M | $135.00 | $945.00 |
The first column is this vendor's lowest-cost tier; the second is the model closest to its median rate, which is nearer what a real deployment bills. Swap in your own token split in the calculator.
Same budget, other vendors
3 models from other vendors priced within 35% of
MiniMax-M2 on a blended rate — the cheapest, the dearest and one in between.
This is the horizontal check a single-vendor pricing page cannot give you.
| Model | Vendor | In / out per 1M | vs this |
|---|---|---|---|
qwen-mt-plus | Alibaba | $0.9 / $2.7 | −31.4% |
gemini-3.5-flash | $0.75 / $4.5 | same | |
gpt-5.6-terra-xhigh | OpenAI | $1 / $6 | +33.3% |
Blended comparison (input + output per 1M tokens). Price is one axis — these are the models worth testing side by side at this budget, not interchangeable substitutes.
Why the output rate decides the bill
Across the 6 priced MiniMax text models, output bills at the same 4.0× the input rate on every one of them. Output tokens are what a chat or agent turn produces — on a typical workload they are the smaller half of the token count but the larger half of the bill.
Worked out: at a 4.0× multiple, a job whose token count is 80% input still spends 50% of its cost on the output it generates. Comparing vendors on the input rate alone hides that.
What MiniMax is good at
MiniMax is competitive for high-volume chat and character-style applications.
Detailed pages for MiniMax models
The models we cover in depth, each with worked cost at three usage levels and a comparison against similarly priced alternatives.
- MiniMax M3 — $0.15 / $0.6 per 1M tokens
Other vendors
Every rate card here is built the same way, from the same price snapshot (2026-09-16), so the numbers are comparable across vendors:
Anthropic · OpenAI · Google · DeepSeek · Alibaba · xAI · Zhipu · Moonshot · Meta · ByteDance
FAQ
How much does MiniMax charge per million tokens?
Across the 6 priced text models in this catalogue, the blended rate (input plus output per 1M tokens) runs from $0.75 on MiniMax-M2.5 to $10.5 on MiniMax-M2.7-highspeed, as of 2026-09-16. The spread is what matters: choosing the wrong tier inside one vendor usually costs more than switching vendor.
What is the cheapest MiniMax model for high-volume work?
MiniMax-M2.5 is the lowest blended rate here at $0.15 input and $0.6 output per 1M tokens. On a Heavy workload — 500M input and 100M output tokens a month — that bills about $135.00. At the other end, the same workload on MiniMax-M2.7-highspeed costs about $1,890.00.
Is MiniMax cheaper through a gateway than buying direct?
We do not claim a discount on MiniMax. Its list pricing is not published in a stable machine-readable form, so every figure on this page is a gateway rate with no verified official price to compare against. Check the number in your dashboard before committing to a budget.
How does output pricing change a MiniMax bill?
Output is billed at a multiple of input on every model priced here, so the input rate alone never predicts the bill. A workload that is mostly input tokens still lands most of its cost on the output column once that multiple is applied. Price your own in/out split in the calculator rather than extrapolating from the headline input rate.
MiniMax only lists a handful of models — does tier choice still matter?
Yes, and more than the model count suggests: the blended spread inside the family is wide enough that picking the wrong tier roughly doubles a monthly bill. With a small catalogue the cheapest way to cut cost is to move a workload down a tier, not to switch vendor.
See all 827 models → or use the cost calculator with your own token split.