Kimi K3 vs Gemma 4 12B: Release Comparison
Comparison of the Kimi K3 and Gemma 4 12B releases: 2 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 4 models.
- For intelligence, Kimi K3 scores highest: Kimi K3 (max) at 44, against Gemma 4 12B (Reasoning) at 14 for Gemma 4 12B.
- For output speed, Gemma 4 12B is fastest: Gemma 4 12B (Reasoning) at 113 t/s, against Kimi K3 (max) at 41 t/s for Kimi K3.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Kimi K3 (max) | 44 | 47 | 50 | 49 | 45 | 43 | 53 | $2.00 | $2.3 | $3.00 | $15.00 | $0.30 | $3,658 | 48k | 32k | 161M | 41 | 4.08s | 53.30s | 65.61s | 1061.60s | 2.8T 104B active at inference time | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | +10 | ||||
| Kimi K3 (low) | 34 | - | - | - | - | - | - | - | $2.3 | $3.00 | $15.00 | $0.30 | - | - | - | - | 37 | 4.72s | 58.10s | 71.44s | - | 2.8T 104B active at inference time | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | |||||
| Gemma 4 12B (Reasoning) | 14 | - | - | - | - | - | - | - | $0.1 | $0.10 | $0.30 | - | - | - | - | - | 113 | 2.37s | 20.12s | 24.55s | - | 12B | 256k | Jun 2026 | - | Yes | Supports: text, image, speech, and video | Supports: text | |||||
| Gemma 4 12B (Non-reasoning) | 9 | - | - | - | - | - | - | - | $0.1 | $0.10 | $0.30 | - | - | - | - | - | 112 | 2.37s | 2.37s | 6.82s | - | 12B | 262k | Jun 2026 | - | No | Supports: text, image, speech, and video | Supports: text | |||||