Artificial Analysis AIトレンド

AIの現状と進歩を牽引する主要トレンドを分析。モデルの知能、効率、アーキテクチャ、推論速度とコスト、学習のトレンドを取り上げます。

AIの進歩

AIの継続的な進歩と、主要AI企業の位置付けを追跡します。

Frontier Language Model Intelligence, Over Time

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

大手テクノロジー企業の設備投資額の推移

四半期ごとの設備投資額(10億USD)

Intelligence Indexとリリース日

Artificial Analysis Intelligence Index · Release date
Most attractive region

AIラボ別の主要モデル

各AIラボが達成したArtificial Analysis Intelligence Indexの最高値

効率

AIの効率がどのように向上しているかを分析します。特定の知能水準を実現するコストと、その知能を利用できる速度の変化も検討します。

Intelligence Index帯別の言語モデル推論料金の推移

100万トークンあたりの料金(USD。キャッシュ、入力、出力トークン料金を7:2:1でブレンド)。帯の区分にはArtificial Analysis Intelligence Index v4.3を使用。

Intelligence Index帯別の言語モデル出力速度の推移

1秒あたりの出力トークン数。帯の区分にはArtificial Analysis Intelligence Index v4.3を使用。

国別分析

AIは世界的な現象です。AIの進歩がどこで起きているか、各国の主要モデルをどう比較できるかについて分析します。AI開発の二大拠点である米国と中国を詳しく取り上げます。

国別の最先端言語モデルの知能の推移

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

オープンウェイト:国別の最先端言語モデルの知能の推移

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Leading Models by Country

Artificial Analysis Intelligence Index · Leading models

オープンソースモデル

オープンウェイトモデルには、柔軟にデプロイでき、特定のユースケースに合わせてファインチューニングできる利点があります。主要なオープンウェイトモデルと、独自モデルとの知能の違いを分析します。

Progress in Open Weights vs. Proprietary Intelligence

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Artificial Analysis Intelligence Index by Open Weights / Proprietary

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

モデルアーキテクチャ

AIモデルのアーキテクチャは、性能と効率に影響します。このセクションでは、Mixture of Experts(MoE)アーキテクチャの普及など、モデルのアーキテクチャに関するトレンドと、モデルの能力との関係を検証します。

モデルアーキテクチャ別のIntelligence Indexとリリース日

Artificial Analysis Intelligence Index、オープンウェイトモデルのみ
Most attractive region

Model Size: Total and Active Parameters

Comparison between total model parameters and parameters active during inference

Intelligence Index vs. Active Parameters

Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line

Intelligence Indexと総パラメーター数

Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line

コンテキスト長(トークン)の四半期中央値

コンテキスト長の中央値(千トークン)

学習分析

AIモデルの学習に関するトレンドとして、学習実行の規模の変化と、学習規模とモデルの知能の関係を分析します。

モデル別の学習トークン

学習トークン数(兆)
No data available

Intelligence Indexと学習トークン数

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1 · 学習トークン数(兆)
Most attractive quadrant
No data available