MiniMax: MiniMax M2
by MiniMaxMiniMax: MiniMax M2 is a large language model from MiniMax. It costs $0.255 per million input tokens and $1.02 per million output tokens. Its context window is 205K tokens.
- Input / 1M tokens
- $0.255
- Output / 1M tokens
- $1.02
- Cached input / 1M
- -
- Context window
- 205K
Not supported
tokens
Who serves it cheapest
3 hosts serve MiniMax M2. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| MinimaxCheapestfp8 | $0.255 | $1.02 | 205K | - | 100.0% |
| $0.300 | $1.20 | 197K | - | 100.0% | |
| Novitafp8 | $0.300 | $1.20 | 205K | - | 100.0% |
The spread between Minimax and Novita is 1.2× for identical weights. Quantization and context limits differ, so check both columns before switching.
About MiniMax M2
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
Specifications
| Model ID | minimax/minimax-m2 |
|---|---|
| Provider | MiniMax |
| Context window | 205K tokens |
| Max output | 131K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - MiniMaxAI/MiniMax-M2 |
| Released | October 23, 2025 |
Cheaper alternatives
Models that cost less than MiniMax M2 while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does MiniMax: MiniMax M2 cost?
$0.255 per million input tokens and $1.02 per million output tokens.
What is the context window of MiniMax: MiniMax M2?
205K tokens, with up to 131K tokens of output per request.
Which provider serves MiniMax: MiniMax M2 cheapest?
Minimax at $0.255 per million input tokens - 1.2× cheaper than Novita, the most expensive of the 3 hosts serving it.
Confirm against the source: MiniMax official pricing.