GLM-5.2 vs Mistral Small 4: Release Comparison
Comparison of the GLM-5.2 and Mistral Small 4 releases: 2 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 4 models.
- For intelligence, GLM-5.2 scores highest: GLM-5.2 (max) at 34, against Mistral Small 4 (Reasoning) at 11 for Mistral Small 4.
- For output speed, Mistral Small 4 is fastest: Mistral Small 4 (Reasoning) at 163 t/s, against GLM-5.2 (Non-reasoning) at 148 t/s for GLM-5.2.
- For cost per task, Mistral Small 4 is cheapest: Mistral Small 4 (Reasoning) at $0.05, against GLM-5.2 (max) at $0.96 for GLM-5.2.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indices
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| GLM-5.2 (max) | 34 | 34 | 31 | 35 | 33 | 37 | 46 | $0.96 | $0.9 | $1.40 | $4.40 | $0.26 | $1,559 | 64k | 51k | 178M | 72 | 8.15s | 35.80s | 42.71s | 839.88s | 753B 40B active at inference time | 1M | Jun 2026 | - | Yes | Supports: text | Supports: text | +16 | ||||
| GLM-5.2 (Non-reasoning) | 22 | - | - | - | - | - | - | - | $0.9 | $1.40 | $4.40 | $0.26 | - | - | - | - | 148 | 1.76s | 1.76s | 5.14s | - | 753B 40B active at inference time | 1M | Jun 2026 | - | No | Supports: text | Supports: text | +4 | ||||
| Mistral Small 4 (Reasoning) | 11 | 11 | 7 | 13 | - | 17 | 18 | $0.05 | $0.2 | $0.15 | $0.60 | - | $79 | 14k | 9k | 54M | 163 | 0.77s | 13.01s | 16.07s | 96.08s | 119B 6.5B active at inference time | 256k | Mar 2026 | - | Yes | Supports: text and image | Supports: text | |||||
| Mistral Small 4 (Non-reasoning) | 9 | - | - | - | - | - | - | - | $0.2 | $0.15 | $0.60 | - | - | - | - | - | 152 | 0.75s | 0.75s | 4.03s | - | 119B 6.5B active at inference time | 256k | Mar 2026 | - | No | Supports: text and image | Supports: text | |||||