Xiaomi releases MiMo-V2-Pro with 1M context window and 1T+ parameters
Xiaomi released MiMo-V2-Pro on March 18, 2026, a flagship foundation model with over 1 trillion total parameters and a 1,048,576 token context window. The model is priced at $1 per million input tokens and $3 per million output tokens, positioning it as an agent-focused system comparable to top-tier models.
MiMo-V2-Pro — Quick Specs
Xiaomi Releases MiMo-V2-Pro With 1M Context Window and 1T+ Parameters
Xiaomi released MiMo-V2-Pro on March 18, 2026, a foundation model featuring over 1 trillion total parameters and a 1,048,576 token context window. The model is priced at $1 per million input tokens and $3 per million output tokens.
Technical Specifications
MiMo-V2-Pro is positioned as Xiaomi's flagship foundation model, designed primarily for agentic scenarios and complex workflow orchestration. The model features:
- Context window: 1,048,576 tokens (1M)
- Parameter count: Over 1 trillion total parameters
- Input pricing: $1 per million tokens
- Output pricing: $3 per million tokens
- Release date: March 18, 2026
Benchmarks and Performance Claims
According to Xiaomi, MiMo-V2-Pro "ranks among the global top tier in the standard PinchBench and ClawBench benchmarks, with perceived performance approaching that of Opus 4.6." The company does not publish specific benchmark scores, instead relying on comparative claims against existing models.
OpenRouter's usage data shows the model handling 319 billion prompt tokens, 1.03 billion completion tokens, and 476 million reasoning tokens in recent tracking periods.
Agentic Focus and Integration
MiMo-V2-Pro is explicitly optimized for agent frameworks, with Xiaomi citing OpenClaw compatibility as a key feature. The company describes the model as "designed to serve as the brain of agent systems, orchestrating complex workflows, driving production engineering tasks, and delivering results reliably."
The model supports reasoning-enabled inference through OpenRouter, allowing access to step-by-step thinking processes via the reasoning parameter in API requests.
Availability and Distribution
MiMo-V2-Pro is accessible through OpenRouter, which handles provider routing and fallback mechanisms to maximize uptime. OpenRouter normalizes API requests and responses across multiple provider implementations.
What This Means
Xiaomi enters the large foundation model market with competitive context window sizing (matching or exceeding most current offerings) and aggressive pricing on the output tier ($3/M output tokens). The focus on agent-optimized design suggests Xiaomi is targeting enterprise automation workflows rather than general-purpose chat applications. The 1M context window places MiMo-V2-Pro in the extended-context category used by models handling document analysis and multi-turn agent reasoning, though specific benchmark data remains unavailable for direct performance comparison. Xiaomi's entry signals continued fragmentation in the foundation model market, with regional players (Chinese tech companies like Xiaomi, Alibaba/Qwen, Baidu, Tencent) establishing independent model lines alongside US-based competitors.
Related Articles
OpenAI Scraps Release of GPT-6.1 Astra Over Safety Concerns
OpenAI confirmed it will not release GPT-6.1 Astra after the model failed to meet internal safety and alignment standards. The decision follows renewed industry-wide calls, including from Anthropic, to slow the pace of frontier model development.
Anthropic Releases Claude Sonnet 5.5, Now Powering Free Tier on Claude.ai
Anthropic released Claude Sonnet 5.5, claiming it runs 30%+ faster and costs up to 30% less than Sonnet 5 while beating it on benchmarks, at the same price. The model now powers the free tier on claude.ai, giving Anthropic a notably stronger free offering than OpenAI's ChatGPT.
Anthropic's Claude Sonnet 5.5 Launches on Amazon Bedrock and Claude Platform on AWS
Anthropic's Claude Sonnet 5.5 is now available on Amazon Bedrock and Claude Platform on AWS, positioned as a faster, lower-cost model for well-scoped coding and document tasks. It pairs with the recently released Claude Opus 5.5, which handles higher-judgment work.
Anthropic Releases Claude Sonnet 5.5: 30% Faster, 30% Cheaper Than Sonnet 5
Anthropic has released Claude Sonnet 5.5, the second model in its Claude 5.5 family following last week's Opus 5.5. The model runs more than 30% faster and costs up to 30% less for most work while keeping Sonnet 5's per-token pricing.
Comments
Loading...