MiniMax-M3

MiniMax🇨🇳 China
active

MiniMax-M3 — frontier open-weight model, downloadable but far beyond consumer hardware.

Context window1000K tokens
Input / 1M tokens$0.3
Output / 1M tokens$1.2
Parameters427B

Version History

1.0major

Initial release of M3 with 428B parameters, native multimodal training, and MiniMax Sparse Attention enabling 1M context with 15× decode speedup over M2.

m3major

M3 introduces MiniMax Sparse Attention to enable 1M-token context at approximately 1/20th the compute cost of previous generation. Native multimodal training on interleaved data with interactive user-simulator tuning.

Benchmark Scores

Full leaderboard →
1445.0 elo
Arena Elo
92.9%
GPQA
18.4%
Hallucination Rate
78.1%
MMMU
80.0 tokens_per_sec
Speed (tok/s)
80.5%
SWE-bench Verified

Coverage