model releaseXiaomi

Xiaomi Launches MiMo-V2.5-Pro with 1M Context Window for Complex Agentic Tasks

TL;DR

Xiaomi released MiMo-V2.5-Pro on April 22, 2026, its flagship model featuring a 1,048,576 token context window and pricing at $1 per million input tokens and $3 per million output tokens. According to Xiaomi, the model ranks highly on ClawEval, GDPVal, and SWE-bench Pro benchmarks, designed for autonomous completion of professional tasks requiring thousands of tool calls.

2 min read
0

MiMo-V2.5-Pro — Quick Specs

Context window1000K tokens
Input$0.435/1M tokens
Output$0.87/1M tokens

Xiaomi Launches MiMo-V2.5-Pro with 1M Context Window for Complex Agentic Tasks

Xiaomi released MiMo-V2.5-Pro on April 22, 2026, its flagship model featuring a 1,048,576 token context window and pricing at $1 per million input tokens and $3 per million output tokens.

Technical Specifications

The model's extended context window — just over 1 million tokens — positions it for integration with agent frameworks requiring long-form context retention. Xiaomi claims the model can autonomously complete professional tasks that would take human experts days or weeks, involving more than a thousand tool calls per task.

Benchmark Performance

According to Xiaomi, MiMo-V2.5-Pro achieves top rankings on:

  • ClawEval: Benchmark for agentic capabilities (specific score not disclosed)
  • GDPVal: General development proficiency evaluation (specific score not disclosed)
  • SWE-bench Pro: Software engineering benchmark (specific score not disclosed)

Xiaomi has not released exact numerical scores for these benchmarks at launch.

Target Use Cases

The company positions MiMo-V2.5-Pro for three primary applications:

  1. General agentic capabilities with extended autonomous operation
  2. Complex software engineering tasks
  3. Long-horizon tasks requiring persistent context across multiple steps

The model is available through OpenRouter, which routes requests to providers with automatic fallbacks for uptime optimization.

Pricing and Availability

MiMo-V2.5-Pro is now available at:

  • Input: $1.00 per million tokens
  • Output: $3.00 per million tokens

This pricing places it in the mid-tier range for flagship models, notably below comparable offerings from Anthropic and OpenAI with similar context windows.

What This Means

Xiaomi's entry with a 1M context window model at competitive pricing adds another option in the expanding agentic AI market. The emphasis on "thousands of tool calls" suggests optimization for complex multi-step workflows rather than single-turn generation. However, without published benchmark scores or independent verification, the claimed performance advantages over existing models remain unconfirmed. The model's actual differentiation will depend on real-world testing in software engineering and agent deployment scenarios.

Related Articles

model release

OpenAI's GPT-6 Astra Reportedly Automates AI Engineering Tasks at Under $6 an Hour, According to Latent Space Testing

A Latent Space report describes GPT-6 Astra, a new OpenAI model the blog says can autonomously handle AI engineering tasks—training models, labeling data, deploying systems—at an estimated cost of under $6 per hour. The claims, including 97.6% on FrontierMath and 99.9% on ARC-AGI-3, come from independent blog testing rather than an official OpenAI announcement.

model release

OpenAI Launches GPT-6 Astra, Says the Model May Already Qualify as AGI

OpenAI has released GPT-6 Astra, its most capable model yet, with benchmark scores the company says surpass GPT-5.6 Sol and Anthropic's Fable 5 models. President Greg Brockman called it a step into the 'AGI era,' though OpenAI acknowledges there's no agreed-upon threshold for that term.

model release

Meta's Muse Spark 1.3 Claims #3 Global Ranking, Matches OpenAI's GPT-5.6-Sol on Coding Benchmarks

Meta Superintelligence Labs shipped Muse Spark 1.3, which the company claims ranks #3 globally on the Artificial Analysis Intelligence Index and matches OpenAI's GPT-5.6-Sol on coding and agentic benchmarks. The model is available now via Muse Code and Meta's API, with open weights and a follow-up model promised soon.

model release

OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold

OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.

Comments

Loading...