MiMo-V2.5 is a next-generation native multimodal AI model developed by Xiaomi MiMo, designed for AI Agents, multimodal reasoning, and complex real-world applications.
The model supports unified understanding across text, images, videos, and audio, enabling advanced interactions across multiple data types. Built with a sparse Mixture-of-Experts (MoE) architecture, MiMo-V2.5 features approximately 310B total parameters with around 15B active parameters, delivering strong performance with improved computational efficiency.
With support for up to 1M-token context windows, MiMo-V2.5 can process long documents, multimodal inputs, and complex workflows. It also includes advanced reasoning, tool calling, web search, and structured output capabilities for building intelligent AI applications.
As part of Xiaomi’s open-source AI model ecosystem, MiMo-V2.5 is released under the MIT License, allowing developers and organizations to customize, deploy, and build commercial AI solutions.
Comments (0)