March 20, 2026
MiMo-V2-Pro: Everything you need to know
See model pageXiaomi has released MiMo-V2-Pro, which scores 49 on the Artificial Analysis Intelligence Index, placing it between Kimi K2.5 and GLM-5
Xiaomi's MiMo-V2-Pro is a new reasoning model and a significant upgrade over their prior open weights release, MiMo-V2-Flash (309B total / 15B active, MIT license), which scores 41 on the Intelligence Index. Xiaomi has not yet released the weights of this model and it is currently only available via Xiaomi's first-party API.
Key takeaways:
-
MiMo-V2-Pro scores 49 on the Artificial Analysis Intelligence Index behind GLM-5 (Reasoning, 50). It is ahead of Kimi K2.5 (Reasoning, 47) and Qwen3.5 397B A17B (Reasoning, 45). On the overall leaderboard, it places #10, just behind GPT-5.2 Codex (xhigh, 49) and ahead of Grok 4.20 Beta (Reasoning, 48)
-
Leading Elo of 1426 on GDPval-AA (Agentic Real-World Work Tasks), ahead of peer models: On GDPval-AA, MiMo-V2-Pro places ahead of GLM-5 (Reasoning, 1406), Kimi K2.5 (Reasoning, 1283), and Qwen3.5 397B A17B (Reasoning, 1209). GPT-5.4 (xhigh) and Claude Sonnet 4.6 (Adaptive Reasoning, max effort) have an Elo of 1667 and 1633 respectively
-
Competitive AA-Omniscience Index driven by low hallucination: MiMo-V2-Pro scores +5, ahead of GLM-5 (Reasoning, +2), Kimi K2.5 (Reasoning, -8), and Qwen3.5 397B A17B (Reasoning, -30). For context, Claude Opus 4.6 (Adaptive Reasoning, max effort, +14) and Gemini 3.1 Pro Preview (+33) remain ahead
-
MiMo-V2-Pro is more token efficient than peers. It used 77M output tokens to run the Artificial Analysis Intelligence Index, significantly less than GLM-5 (Reasoning, 109M) and Kimi K2.5 (Reasoning, 89M)
-
MiMo-V2-Pro costs $348 to run the Artificial Analysis Intelligence Index at $1/$3 per 1M input/output tokens. This is less expensive than GLM-5 despite scoring only 1 point lower on the Intelligence Index. For comparison, GPT-5.2 (xhigh) cost $2,304 and Claude Opus 4.6 (Adaptive Reasoning, max effort) cost $2,486
Key model information:
➤ Context window: 1M tokens
➤ Pricing: $1/$3 per 1M input/output tokens, for 256K token input and $2/$6 per 1M input/output tokens for 1M token input
➤ Availability: Xiaomi first-party API only
➤ Modality: Text input and output only (no multimodality)

MiMo-V2-Pro has an Elo of 1426 for GDPval-AA, ahead of other models from China

MiMo-V2-Pro cost $348 to run the Artificial Analysis Intelligence Index at $1/$3 per 1M input/output tokens and is on the Pareto frontier of the Intelligence Index vs. Cost to Run Intelligence Index chart

MiMo-V2-Pro generated ~70M reasoning tokens while running the Artificial Analysis Intelligence Index, less than select peers of equivalent intelligence

MiMo-V2-Pro scores 5 on the AA-Omniscience Index, driven primarily by a low hallucination rate of 30%. This is an improvement over MiMo-V2-Flash’s 48% hallucination rate

Full breakdown of results by individual evaluations

See Artificial Analysis for further details and benchmarks of MiMo-V2-Pro: https://artificialanalysis.ai/models/mimo-v2-pro
Read the latest

Agnes AI releases Agnes 2.5 Pro Beta
Agnes 2.5 Pro Beta
August 27, 2026

Intelligence at pocket scale: Benchmarking small models and mobile phones
Independent intelligence benchmarking of small language models on a set of evaluations chosen for mobile device use, launched alongside mobile phone inference benchmarking with Liquid AI. We evaluate the same quantized builds used on mobile phones, and performance is measured on real devices.
August 24, 2026

Announcing the Speech Agent Arena: Compare Speech agents in real world conversations
Announcing our new Speech Agent Arena, evaluating Speech to Speech models on real-world scenarios to analyze conversational preference and task success rate
August 24, 2026