Wafer:モデルの知能、性能、料金

Wafer
Wafer

この分析は、ユースケースに最適なWafer提供モデルを選ぶための参考情報です。

最高の知能

Updated
#1
GLM-5.3 (max)GLM-5.3 (max)
45
#2
GLM-5.3-FlashGLM-5.3-Flash
42
#3
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
39
#4
DeepSeek V4 Flash 0731 (max) (FAST)DeepSeek V4 Flash 0731 (max) (FAST)
34
#5
Qwen3.8 27B (xhigh)Qwen3.8 27B (xhigh)
34

Intelligence Index

モデル合計:6件

最速

#1
GLM-5.3 (max)GLM-5.3 (max)
180 t/s
#2
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
142 t/s
#3
DeepSeek V4 Flash 0731 (max) (FAST)DeepSeek V4 Flash 0731 (max) (FAST)
119 t/s
#4
Qwen3.8 27B (non-reasoning)Qwen3.8 27B (non-reasoning)
81 t/s
#5
Qwen3.8 27B (xhigh)Qwen3.8 27B (xhigh)
77 t/s

出力速度

モデル合計:6件

最安料金

#1
GLM-5.3-FlashGLM-5.3-Flash
$0.10
#2
DeepSeek V4 Flash 0731 (max) (FAST)DeepSeek V4 Flash 0731 (max) (FAST)
$0.16
#3
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
$0.18
#4
Qwen3.8 27B (xhigh)Qwen3.8 27B (xhigh)
$0.31
#5
Qwen3.8 27B (non-reasoning)Qwen3.8 27B (non-reasoning)
$0.31

ブレンド料金(100万トークンあたり)

モデル合計:6件

Waferは6モデルを提供しており、知能、性能、料金の特性はそれぞれ異なります。 以下でモデル間の主要指標を比較します。

  • Waferで知能が上位のモデルはGLM-5.3 (max)(45)、GLM-5.3-Flash(42)、DeepSeek V4.1 Flash (max)(39)です。
  • 出力速度が最も速いモデルはGLM-5.3 (max)(180 t/s)、DeepSeek V4.1 Flash (max)(142 t/s)、DeepSeek V4 Flash 0731 (max) (FAST)(119 t/s)です。 モデル間の速度差は大きく、最速と最遅で134%の差があります。
  • 遅延では、Qwen3.8 27B (non-reasoning)(1.26秒)、GLM-5.3 (max)(11.95秒)、DeepSeek V4.1 Flash (max)(14.82秒)の最初の回答トークンまでの時間が最短です。
  • 料金では、GLM-5.3-Flash($0.10)、DeepSeek V4 Flash 0731 (max) (FAST)($0.16)、DeepSeek V4.1 Flash (max)($0.18)の100万トークンあたりのブレンド料金が最安です。 モデル間で料金に最大3.1倍の差があります。
  • Waferで最大のコンテキストウィンドウに対応するモデルはDeepSeek V4.1 Flash (max)(1M)、GLM-5.3 (max)(1M)、GLM-5.3-Flash(1M)です。
  • GLM-5.3 (max)は知能と速度を最も高い水準で両立しています。費用を最適化するなら、GLM-5.3-Flashの料金が最も競争力に優れています。
Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

知能評価

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
もっと見る

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

SciCodeUnder review

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

No data available

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

No data available

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Indexと料金

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

コンテキストウィンドウ

コンテキストウィンドウ

Context window: tokens limit · Higher is better

料金

Intelligence Indexと料金

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

性能の概要

出力速度と料金

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

速度

出力速度(1秒あたりのトークン数)で測定

出力速度

Output tokens per second · Higher is better

遅延

最初のトークンまでの時間(秒)で測定

遅延: 最初の回答トークンまでの時間

最初の回答トークン受信までの秒数 · 推論モデルの「思考」時間を含む

キャッシュ動作

キャッシュヒット率

キャッシュ可能な入力トークンのうちキャッシュから提供された割合 · 直近4週間の中央値、2026年9月26日更新

タスクあたりのコストとキャッシュヒット率

Weighted average cost (USD) per Intelligence Index task · Cache hit rate: median of the last four weeks, updated Sep 26, 2026
Most attractive quadrant

エンドツーエンド応答時間

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

エンドツーエンド応答時間と料金

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

詳細分析
Z AIのロゴ
GLM-5.3 (max)
1M
オープン
45
$2.01
172
0.85
15.40
11.64
Z AIのロゴ
GLM-5.3-Flash
1M
オープン
42
$0.45
64
0.92
39.89
31.18
DeepSeekのロゴ
DeepSeek V4.1 Flash (max)
1.05M
オープン
39
$0.57
152
0.72
17.19
13.18
DeepSeekのロゴ
DeepSeek V4 Flash 0731 (max) (FAST)
1M
オープン
34
$0.54
122
0.77
21.31
16.43
Alibabaのロゴ
Qwen3.8 27B (xhigh)
262k
オープン
34
$0.45
77
1.24
33.66
25.93
Alibabaのロゴ
Qwen3.8 27B (non-reasoning)
262k
オープン
20
$0.92
79
1.26
7.56
--

主要な定義

よくある質問

Waferに関するよくある質問