Xiaomi Launches MiMo-V2.5-Pro with 1M Context Window for Complex Agentic Tasks
Xiaomi released MiMo-V2.5-Pro on April 22, 2026, its flagship model featuring a 1,048,576 token context window and pricing at $1 per million input tokens and $3 per million output tokens. According to Xiaomi, the model ranks highly on ClawEval, GDPVal, and SWE-bench Pro benchmarks, designed for autonomous completion of professional tasks requiring thousands of tool calls.
MiMo-V2.5-Pro — Quick Specs
Xiaomi Launches MiMo-V2.5-Pro with 1M Context Window for Complex Agentic Tasks
Xiaomi released MiMo-V2.5-Pro on April 22, 2026, its flagship model featuring a 1,048,576 token context window and pricing at $1 per million input tokens and $3 per million output tokens.
Technical Specifications
The model's extended context window — just over 1 million tokens — positions it for integration with agent frameworks requiring long-form context retention. Xiaomi claims the model can autonomously complete professional tasks that would take human experts days or weeks, involving more than a thousand tool calls per task.
Benchmark Performance
According to Xiaomi, MiMo-V2.5-Pro achieves top rankings on:
- ClawEval: Benchmark for agentic capabilities (specific score not disclosed)
- GDPVal: General development proficiency evaluation (specific score not disclosed)
- SWE-bench Pro: Software engineering benchmark (specific score not disclosed)
Xiaomi has not released exact numerical scores for these benchmarks at launch.
Target Use Cases
The company positions MiMo-V2.5-Pro for three primary applications:
- General agentic capabilities with extended autonomous operation
- Complex software engineering tasks
- Long-horizon tasks requiring persistent context across multiple steps
The model is available through OpenRouter, which routes requests to providers with automatic fallbacks for uptime optimization.
Pricing and Availability
MiMo-V2.5-Pro is now available at:
- Input: $1.00 per million tokens
- Output: $3.00 per million tokens
This pricing places it in the mid-tier range for flagship models, notably below comparable offerings from Anthropic and OpenAI with similar context windows.
What This Means
Xiaomi's entry with a 1M context window model at competitive pricing adds another option in the expanding agentic AI market. The emphasis on "thousands of tool calls" suggests optimization for complex multi-step workflows rather than single-turn generation. However, without published benchmark scores or independent verification, the claimed performance advantages over existing models remain unconfirmed. The model's actual differentiation will depend on real-world testing in software engineering and agent deployment scenarios.
Related Articles
OpenAI's GPT-6 Astra Reportedly Automates AI Engineering Tasks at Under $6 an Hour, According to Latent Space Testing
A Latent Space report describes GPT-6 Astra, a new OpenAI model the blog says can autonomously handle AI engineering tasks—training models, labeling data, deploying systems—at an estimated cost of under $6 per hour. The claims, including 97.6% on FrontierMath and 99.9% on ARC-AGI-3, come from independent blog testing rather than an official OpenAI announcement.
OpenAI Launches GPT-6 Astra, Says the Model May Already Qualify as AGI
OpenAI has released GPT-6 Astra, its most capable model yet, with benchmark scores the company says surpass GPT-5.6 Sol and Anthropic's Fable 5 models. President Greg Brockman called it a step into the 'AGI era,' though OpenAI acknowledges there's no agreed-upon threshold for that term.
Meta's Muse Spark 1.3 Claims #3 Global Ranking, Matches OpenAI's GPT-5.6-Sol on Coding Benchmarks
Meta Superintelligence Labs shipped Muse Spark 1.3, which the company claims ranks #3 globally on the Artificial Analysis Intelligence Index and matches OpenAI's GPT-5.6-Sol on coding and agentic benchmarks. The model is available now via Muse Code and Meta's API, with open weights and a follow-up model promised soon.
OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold
OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.
Comments
Loading...