Groq:モデルの知能、性能、料金

Groq
Groq

この分析は、ユースケースに最適なGroq提供モデルを選ぶための参考情報です。

最高の知能

Updated
#1
Qwen3.6 27BQwen3.6 27B
22
#2
Qwen3.6 27B (Non-reasoning)Qwen3.6 27B (Non-reasoning)
20
#3
gpt-oss-120b (high)gpt-oss-120b (high)
12
#4
gpt-oss-120b (low)gpt-oss-120b (low)
10
#5
gpt-oss-20b (low)gpt-oss-20b (low)
10

Intelligence Index

モデル合計:8件

最速

#1
gpt-oss-20b (high)gpt-oss-20b (high)
959 t/s
#2
gpt-oss-20b (low)gpt-oss-20b (low)
931 t/s
#3
Llama 3.1 8BLlama 3.1 8B
624 t/s
#4
gpt-oss-120b (high)gpt-oss-120b (high)
474 t/s
#5
gpt-oss-120b (low)gpt-oss-120b (low)
473 t/s

出力速度

モデル合計:8件

最安料金

#1
Llama 3.1 8BLlama 3.1 8B
$0.05
#2
gpt-oss-20b (low)gpt-oss-20b (low)
$0.10
#3
gpt-oss-20b (high)gpt-oss-20b (high)
$0.10
#4
gpt-oss-120b (high)gpt-oss-120b (high)
$0.14
#5
gpt-oss-120b (low)gpt-oss-120b (low)
$0.20

ブレンド料金(100万トークンあたり)

モデル合計:8件

Groqは8モデルを提供しており、知能、性能、料金の特性はそれぞれ異なります。 以下でモデル間の主要指標を比較します。

  • Groqで知能が上位のモデルはQwen3.6 27B(22)、Qwen3.6 27B (Non-reasoning)(20)、gpt-oss-120b (high)(12)です。
  • 出力速度が最も速いモデルはgpt-oss-20b (high)(959 t/s)、gpt-oss-20b (low)(931 t/s)、Llama 3.1 8B(624 t/s)です。 モデル間の速度差は大きく、最速と最遅で103%の差があります。
  • 遅延では、Llama 3.3 70B(0.81秒)、Llama 3.1 8B(0.94秒)、Qwen3.6 27B (Non-reasoning)(1.19秒)の最初の回答トークンまでの時間が最短です。
  • 料金では、Llama 3.1 8B($0.05)、gpt-oss-20b (low)($0.10)、gpt-oss-20b (high)($0.10)の100万トークンあたりのブレンド料金が最安です。 モデル間で料金に最大3.7倍の差があります。
  • Groqで最大のコンテキストウィンドウに対応するモデルはQwen3.6 27B(131k)、Qwen3.6 27B (Non-reasoning)(131k)、gpt-oss-120b (high)(131k)です。

ハイライト

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

知能評価

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

No data available

Instruction following

Agentic tool use

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

コンテキストウィンドウ

Context Window

Context window: tokens limit · Higher is better

料金

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

性能の概要

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

速度

出力速度(1秒あたりのトークン数)で測定

Output Speed

Output tokens per second · Higher is better

遅延

最初のトークンまでの時間(秒)で測定

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

エンドツーエンド応答時間

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant

詳細分析
Alibabaのロゴ
Qwen3.6 27B
131k
オープン
22
$0.60
444
1.22
15.12
12.77
Alibabaのロゴ
Qwen3.6 27B (Non-reasoning)
131k
オープン
20*
--
452
1.19
2.29
--
OpenAIのロゴ
gpt-oss-120b (high)
131k
オープン
12
$0.08
474
0.72
5.99
4.22
OpenAIのロゴ
gpt-oss-120b (low)
131k
オープン
10*
--
473
0.69
5.98
4.23
OpenAIのロゴ
gpt-oss-20b (low)
131k
オープン
10*
--
931
0.84
3.53
2.15
OpenAIのロゴ
gpt-oss-20b (high)
131k
オープン
9
$0.02
959
0.78
3.39
2.09
Metaのロゴ
Llama 3.3 70B
131k
オープン
8*
--
299
0.81
2.49
--
Metaのロゴ
Llama 3.1 8B
131k
オープン
7*
--
624
0.94
1.74
--

主要な定義

よくある質問

Groqに関するよくある質問