MiMo-V2-Flash vs Claude 4.5 Haiku: Release Comparison
Comparison of the MiMo-V2-Flash and Claude 4.5 Haiku releases: 2 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 4 models.
- For intelligence, MiMo-V2-Flash scores highest: MiMo-V2-Flash (Reasoning) at 32, against Claude 4.5 Haiku (Reasoning) at 30 for Claude 4.5 Haiku.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indices
Measures performance in agentic workflows, focusing on behaviors like tool use, planning, autonomy, and complex problem solving.
Incorporates 5 evaluations: AA-Omniscience, GDPval-AA v2, Humanity's Last Exam, 𝜏³-Banking, AA-LCR · Higher is better
Incorporates 4 evaluations: AA-Omniscience, GDPval-AA v2, 𝜏³-Banking, AA-LCR · Higher is better
Incorporates 5 evaluations: AA-Omniscience, GDPval-AA v2, AA-LCR, 𝜏³-Banking, Humanity's Last Exam · Higher is better
Incorporates 5 evaluations: AA-Omniscience, GDPval-AA v2, MLCR-AA, Humanity's Last Exam, 𝜏³-Banking · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, GPQA Diamond, CritPt, GDPval-AA v2, Terminal-Bench v2.1 · Higher is better
Incorporates 4 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-LCR · Higher is better
Further details
Weights | Provider Benchmarks | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| MiMo-V2-Flash (Reasoning) | 32 | 309B 15B active at inference time | 256k | $0.1 | - | ||||
| Claude 4.5 Haiku (Reasoning) | 30 | - | 200k | $0.8 | 92 | Not available | |||
| MiMo-V2-Flash (Non-reasoning) | 25 | 309B 15B active at inference time | 256k | - | - | - | |||
| Claude 4.5 Haiku (Non-reasoning) | 24 | - | 200k | $0.8 | 85 | Not available |