MiMo-V2.5-Pro

3.95
MiMo-V2.5-Pro is a flagship MoE model specialized in long-term agent tasks and complex coding. It stably executes over 1,000 tool calls, reduces token consumption by 40–60% vs. competitors, and is open-sourced under MIT license.
Advertisement 728 × 90
CompanyXiaomi
Context1M
Released2026-04
Updated2026-07-20

MiMo-V2.5-Pro Overview

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro. It can independently and autonomously complete professional tasks that would take human experts days or weeks, involving more than a thousand tool calls. Its context length of up to 1M makes it well suited for integration with a wide range of agent frameworks.

MiMo-V2.5-Pro Pricing

PlanPriceDescription
Xiaomi API Access Input $0.435 / Output $0.87 per 1M Tokens Supports up to a 1M-token context window and up to 128K maximum output tokens. Built on an approximately 1T-parameter MoE architecture with around 42B active parameters. Cached input pricing is approximately $0.0036 per 1M Tokens.
UltraSpeed Mode Input $1.305 / Output $2.61 per 1M Tokens A high-speed inference mode designed for latency-sensitive applications. Offers faster response performance at approximately 3× the standard API price.
OpenRouter API Access Potentially lower than standard API pricing Access MiMo-V2.5 Pro through OpenRouter and other third-party providers. Some providers may offer promotional pricing or discounts. The model is released under the MIT license.

MiMo-V2.5-Pro Key Features

1. Superior Agentic Capabilities: Optimized for complex agent tasks, scoring 72.9 on τ³-bench, on par with GPT-5.4
2. Excellent Code Generation: Scores 73.7 on MiMo Coding Bench; developed a full compiler in 4.3 hours
3. Efficient MoE Architecture: 1.02T total params, only 42B activated per inference
4. Million-Token Context Window: 1M context for long documents and complex multi-turn tasks
5. Open-Source under MIT: Fully open-source, free for commercial deployment and fine-tuning
6. Ultra-Fast Inference: UltraSpeed version achieves 1000+ tokens/second
7. Token Efficiency: Saves 40%-60% token consumption compared to competitors

Summary

1.AI Agent Developers: Those building complex multi-step automation workflows and agentic applications
2.Software Engineers: Developers needing code generation, code review, and project building assistance
3.Enterprise Users: Organizations requiring long-document processing and complex task automation
4.Open-Source Community & Researchers: Individuals and teams leveraging high-performance open models for research and development
5.AI Product Teams: Teams needing rapid prototyping and high-performance inference

Comments (0)

Leave a comment

Advertisement 728 × 90

Xiaomi Model Comparison

Model Context Pricing API Released Global Heat
MiMo-V2.5-Pro
1M Paid YES 2026-04
79/100
1.1M YES 2026-04
52/100

Similar Models

Related News