StepFun:モデルの知能、性能、料金

StepFun
StepFun

この分析は、ユースケースに最適なStepFun提供モデルを選ぶための参考情報です。

最高の知能

Updated
#1
Step 3.7 FlashStep 3.7 Flash
19
#2
Step 3.5 Flash 2603Step 3.5 Flash 2603
17
#3
Step 3.5 FlashStep 3.5 Flash
17

Intelligence Index

モデル合計:3件

最速

#1
Step 3.5 FlashStep 3.5 Flash
151 t/s
#2
Step 3.5 Flash 2603Step 3.5 Flash 2603
146 t/s
#3
Step 3.7 FlashStep 3.7 Flash
113 t/s

出力速度

モデル合計:3件

最安料金

#1
Step 3.5 Flash 2603Step 3.5 Flash 2603
$0.06
#2
Step 3.5 FlashStep 3.5 Flash
$0.12
#3
Step 3.7 FlashStep 3.7 Flash
$0.18

ブレンド料金(100万トークンあたり)

モデル合計:3件

StepFunは3モデルを提供しており、知能、性能、料金の特性はそれぞれ異なります。 以下でモデル間の主要指標を比較します。

  • StepFunで知能が上位のモデルはStep 3.7 Flash(19)、Step 3.5 Flash 2603(17)、Step 3.5 Flash(17)です。
  • 出力速度が最も速いモデルはStep 3.5 Flash(151 t/s)、Step 3.5 Flash 2603(146 t/s)、Step 3.7 Flash(113 t/s)です。
  • 遅延では、Step 3.5 Flash(16.68秒)、Step 3.5 Flash 2603(17.26秒)、Step 3.7 Flash(20.38秒)の最初の回答トークンまでの時間が最短です。
  • 料金では、Step 3.5 Flash 2603($0.06)、Step 3.5 Flash($0.12)、Step 3.7 Flash($0.18)の100万トークンあたりのブレンド料金が最安です。
  • StepFunで最大のコンテキストウィンドウに対応するモデルはStep 3.7 Flash(256k)、Step 3.5 Flash 2603(256k)、Step 3.5 Flash(256k)です。

ハイライト

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

知能評価

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

No data available

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

No data available

Agentic coding & terminal use

No data available

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

No data available

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

No data available

Instruction following

Agentic tool use

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

コンテキストウィンドウ

Context Window

Context window: tokens limit · Higher is better

料金

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

性能の概要

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

速度

出力速度(1秒あたりのトークン数)で測定

Output Speed

Output tokens per second · Higher is better

遅延

最初のトークンまでの時間(秒)で測定

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

エンドツーエンド応答時間

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

詳細分析
StepFunのロゴ
Step 3.7 Flash
256k
オープン
19*
--
113
2.67
24.80
17.71
StepFunのロゴ
Step 3.5 Flash 2603
256k
独自
17*
--
146
3.55
20.68
13.71
StepFunのロゴ
Step 3.5 Flash
256k
オープン
17*
--
151
3.40
20.00
13.28

主要な定義

よくある質問

StepFunに関するよくある質問