Claude Opus 5 vs. Mistral Small 4:リリース比較
Claude Opus 5とMistral Small 4のリリース比較:それぞれ5個と2個のモデルがあり、知能・性能・料金の特性が異なります。以下では、全7個のモデルの主要指標を比較します。
- 知能が最も高いのはClaude Opus 5です。Claude Opus 5 (Adaptive Reasoning, Max Effort)(51)に対し、Mistral Small 4はMistral Small 4 (Reasoning)(11)です。
- 出力速度が最も速いのはMistral Small 4です。Mistral Small 4 (Reasoning)(159トークン/秒)に対し、Claude Opus 5はClaude Opus 5 (Adaptive Reasoning, Medium Effort)(57トークン/秒)です。
- タスクあたりの費用が最も低いのはMistral Small 4です。Mistral Small 4 (Reasoning)($0.05)に対し、Claude Opus 5はClaude Opus 5 (Adaptive Reasoning, Low Effort)($1.10)です。
比較するリリース
知能
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
費用
Cost per Intelligence Index Task
速度と遅延
Output Speed
能力スコア
能力指数
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
詳細
入力モダリティ | 出力モダリティ | ウェイト | プロバイダーのベンチマーク | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Claude Opus 5 (Adaptive Reasoning, Max Effort) | 51 | 55 | 57 | 57 | 53 | 54 | 61 | $5.86 | $3.9 | $5.00 | $25.00 | $0.50 | $7,275 | 73k | 43k | 140M | 55 | 62.61秒 | 62.61秒 | 71.71秒 | 804.00秒 | - | 1M | 2026年7月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | 利用不可 | - | |||
| Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) | 50 | 54 | 54 | 56 | 53 | 53 | 60 | $4.88 | $3.9 | $5.00 | $25.00 | $0.50 | $5,868 | 61k | 35k | 111M | 54 | 26.58秒 | 26.58秒 | 35.92秒 | 724.44秒 | - | 1M | 2026年7月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | 利用不可 | - | |||
| Claude Opus 5 (Adaptive Reasoning, High Effort) | 48 | 51 | 53 | 54 | 52 | 52 | 58 | $3.61 | $3.9 | $5.00 | $25.00 | $0.50 | $4,332 | 46k | 25k | 81M | 55 | 16.08秒 | 16.08秒 | 25.14秒 | 517.71秒 | - | 1M | 2026年7月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | 利用不可 | - | |||
| Claude Opus 5 (Adaptive Reasoning, Medium Effort) | 45 | 49 | 51 | 52 | 50 | 47 | 56 | $2.19 | $3.9 | $5.00 | $25.00 | $0.50 | $2,732 | 29k | 15k | 49M | 57 | 5.20秒 | 5.20秒 | 13.99秒 | 314.21秒 | - | 1M | 2026年7月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | 利用不可 | - | |||
| Claude Opus 5 (Adaptive Reasoning, Low Effort) | 39 | 44 | 46 | 48 | 45 | 42 | 50 | $1.10 | $3.9 | $5.00 | $25.00 | $0.50 | $1,561 | 15k | 7k | 26M | 56 | 2.94秒 | 2.94秒 | 11.92秒 | 164.81秒 | - | 1M | 2026年7月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | 利用不可 | - | |||
| Mistral Small 4 (Reasoning) | 11 | 10 | 7 | 13 | - | 16 | 18 | $0.05 | $0.2 | $0.15 | $0.60 | - | $79 | 14k | 9k | 54M | 159 | 0.73秒 | 13.27秒 | 16.41秒 | 93.03秒 | 119B 推論時に6.5Bが有効 | 256k | 2026年3月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | |||||
| Mistral Small 4 (Non-reasoning) | 9 | - | - | - | - | - | - | - | $0.2 | $0.15 | $0.60 | - | - | - | - | - | 151 | 0.72秒 | 0.72秒 | 4.03秒 | - | 119B 推論時に6.5Bが有効 | 256k | 2026年3月 | - | いいえ | 対応:テキスト、画像 | 対応:テキスト | |||||