PulseAugur
EN
LIVE 04:32:38

Xiaomi unveils native omnimodal MiMo-V2.5 model with 1M context window

Xiaomi has released its MiMo-V2.5 model, notable for its "native omnimodal" architecture. Unlike many multimodal models that integrate separate components for different data types, MiMo-V2.5 was designed from the ground up as a unified system capable of processing text, images, video, and audio simultaneously. This Mixture-of-Experts model boasts approximately 310 billion total parameters with 15 billion active per token, and supports a context window of up to 1 million tokens. AI

IMPACT This native omnimodal architecture could set a new standard for processing diverse data types, potentially improving agent capabilities and long-context reasoning.

RANK_REASON New model release from a major tech company (Xiaomi) with novel architectural claims ('native omnimodal'). [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on dev.to — LLM tag →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Xiaomi unveils native omnimodal MiMo-V2.5 model with 1M context window

COVERAGE [1]

  1. dev.to — LLM tag TIER_1 English(EN) · RouteAI ·

    MiMo-V2.5 Explained: Xiaomi's Native Omnimodal Model and What "Native" Actually Means

    <p>Xiaomi released MiMo-V2.5 in April 2026, and the detail worth paying attention to isn't the benchmark numbers — it's the word "native" in "native omnimodal."</p> <p><a class="article-body-image-wrapper" href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-…