MiniMax API pricing: every model, every rate
MiniMax ships a compact family with a notable gap between standard and highspeed variants — the latter bills several times more for lower latency.
Full rate card
6 billable models. 0 carry a rate verified against the vendor's own published pricing; the rest are gateway rates only. USD per 1M tokens, verified 2026-09-16.
| Model | Input | Output | Out/In | Official (in / out) |
|---|---|---|---|---|
MiniMax-M2.7-highspeed | $2.1 | $8.4 | 4.0× | not verified |
MiniMax-M2 | $1.05 | $4.2 | 4.0× | not verified |
MiniMax-M2.1 | $1.05 | $4.2 | 4.0× | not verified |
MiniMax-M2.5 | $0.15 | $0.6 | 4.0× | not verified |
MiniMax-M2.7 | $0.15 | $0.6 | 4.0× | not verified |
MiniMax-M3 | $0.15 | $0.6 | 4.0× | not verified |
Output is billed at a multiple of input on most models. The Out/In column is that multiple — it matters more than the input price once output dominates your bill.
Across 6 billable models the rate spans $0.15 (MiniMax-M2.5) to $2.1 (MiniMax-M2.7-highspeed) per million input tokens — a 14× spread. Picking the wrong tier is usually the single most expensive mistake here.
What MiniMax is good at
MiniMax is competitive for high-volume chat and character-style applications.
See all 827 models → or use the cost calculator with your own token split.