MiMo-V2.6-Pro-RL

Xiaomi🇨🇳 China
active
Context window1000K tokens

Version History

V2.6-Pro-RLmajor

MiMo-V2.6-Pro-RL is a new 1.02T-parameter (42B active) omnimodal MoE model trained via a single mixed RL run, showing large benchmark gains over the prior MiMo-V2.5-Pro checkpoint, especially in cybersecurity and agentic tasks.

Coverage

model releaseXiaomi

Xiaomi Releases MiMo-V2.6-Pro-RL, a 1.02T-Parameter Omnimodal Model with 1M-Token Context

Xiaomi's MiMo team has released MiMo-V2.6-Pro-RL, a 1.02-trillion-parameter sparse mixture-of-experts model with 42B active parameters, 1M-token context, and native text/image/video/audio processing. The model was trained via a single mixed reinforcement learning run spanning coding, agentic, visual, and cybersecurity tasks, with benchmark scores that Xiaomi claims approach or match Claude Opus 5 and GPT-5.6 on several agentic and coding tests.

3 min read