Claude Opus 5 vs. Kimi K3:リリース比較

Claude Opus 5とKimi K3のリリース比較:それぞれ5個と2個のモデルがあり、知能・性能・料金の特性が異なります。以下では、全7個のモデルの主要指標を比較します。

  • 知能が最も高いのはClaude Opus 5です。Claude Opus 5 (Adaptive Reasoning, Max Effort)(51)に対し、Kimi K3はKimi K3 (max)(44)です。
  • 出力速度が最も速いのはClaude Opus 5です。Claude Opus 5 (Adaptive Reasoning, Medium Effort)(57トークン/秒)に対し、Kimi K3はKimi K3 (max)(38トークン/秒)です。
  • タスクあたりの費用が最も低いのはClaude Opus 5です。Claude Opus 5 (Adaptive Reasoning, Low Effort)($1.10)に対し、Kimi K3はKimi K3 (max)($2.00)です。

比較するリリース

知能

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Opus 5
Pareto line

Cost per Task (USD, Log Scale)

費用

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

速度と遅延

Output Speed

Output tokens per second · Higher is better

能力スコア

能力指数

特定の能力や業界におけるモデルの性能を測定
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better

詳細

入力モダリティ
出力モダリティ
ウェイト
プロバイダーのベンチマーク
Claude Opus 5 (Adaptive Reasoning, Max Effort)
AnthropicのロゴAnthropic
51
55
57
57
53
54
61
$5.86
$3.9
$5.00
$25.00
$0.50
$7,275
73k
43k
140M
55
62.61秒
62.61秒
71.71秒
804.00秒
-
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

利用不可
-
Amazon BedrockAnthropicGoogle
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
AnthropicのロゴAnthropic
50
54
54
56
53
53
60
$4.88
$3.9
$5.00
$25.00
$0.50
$5,868
61k
35k
111M
54
26.58秒
26.58秒
35.92秒
724.44秒
-
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

利用不可
-
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, High Effort)
AnthropicのロゴAnthropic
48
51
53
54
52
52
58
$3.61
$3.9
$5.00
$25.00
$0.50
$4,332
46k
25k
81M
55
16.08秒
16.08秒
25.14秒
517.71秒
-
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

利用不可
-
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, Medium Effort)
AnthropicのロゴAnthropic
45
49
51
52
50
47
56
$2.19
$3.9
$5.00
$25.00
$0.50
$2,732
29k
15k
49M
57
5.20秒
5.20秒
13.99秒
314.21秒
-
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

利用不可
-
AnthropicAmazon BedrockGoogle
Kimi K3 (max)
KimiのロゴKimi
44
47
50
49
45
43
53
$2.00
$2.3
$3.00
$15.00
$0.30
$3,658
48k
32k
161M
38
4.15秒
56.45秒
69.52秒
1068.17秒
2.8T
推論時に104Bが有効
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

DatabricksFireworksModal
+10
Claude Opus 5 (Adaptive Reasoning, Low Effort)
AnthropicのロゴAnthropic
39
44
46
48
45
42
50
$1.10
$3.9
$5.00
$25.00
$0.50
$1,561
15k
7k
26M
56
2.94秒
2.94秒
11.92秒
164.81秒
-
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

利用不可
-
AnthropicAmazon BedrockGoogle
Kimi K3 (low)
KimiのロゴKimi
34
-
-
-
-
-
-
-
$2.3
$3.00
$15.00
$0.30
-
-
-
-
37
4.72秒
58.49秒
71.93秒
-
2.8T
推論時に104Bが有効
1M
2026年7月
-
はい

対応:テキスト、画像

対応:テキスト

KimiBasetenDatabricks