모든 릴리스•

모델 22개

Meta 모델: 지능, 성능 및 가격

Artificial Analysis는 Meta의 모델 22개를 벤치마크했습니다. 아래에서 이 모델들의 주요 지표를 비교합니다.

  • 지능 부문에서 Meta의 최고 모델은 Muse Spark 1.3 (max)(48)입니다.
  • 출력 속도가 가장 빠른 모델은 Muse Spark 1.3 (xhigh)(221 토큰/초)입니다.
  • 지연 시간 부문에서 첫 응답 토큰까지 걸리는 시간이 가장 짧은 모델은 Llama 4 Scout(0.84초)입니다.
  • 가격 부문에서 작업당 비용이 가장 낮은 모델은 Muse Glimmer (high)($0.06)입니다. 모델 간 가격 차이는 최대 28배입니다.

지능

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Pareto line

비용

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

속도 및 지연 시간

Output Speed

Output tokens per second · Higher is better

역량 점수

역량 지수

특정 역량 및 산업에서 모델의 성능을 측정
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better

Meta의 모든 릴리스

상세 정보

웨이트
제공업체 벤치마크
Muse Spark 1.3 (max)
Meta 로고Meta
48
-
1M
US$0.8
203
제공되지 않음
Meta
Muse Spark 1.3 (xhigh)
Meta 로고Meta
45
-
1M
US$0.8
221
제공되지 않음
Meta
Muse Spark 1.2 (xhigh)
Meta 로고Meta
40
-
1M
US$0.8
234
제공되지 않음
Meta
Muse Spark 1.1 (xhigh)
Meta 로고Meta
34
-
1M
US$0.8
-
제공되지 않음
Meta
Muse Spark
Meta 로고Meta
31
-
262k
-
-
제공되지 않음
-
Muse Glimmer (high)
Meta 로고Meta
17
30B
131k
US$0.2
122
Together AIDeepInfraSystalyzeFireworks
Llama 4 Maverick
Meta 로고Meta
10
402B
추론 시 17B 활성
1M
US$0.3
110
NovitaMicrosoft AzureAmazon Bedrock
+3
Llama 4 Scout
Meta 로고Meta
8
109B
추론 시 17B 활성
10M
US$0.2
107
DeepInfraMicrosoft AzureGoogle
+3
Llama 3.3 Instruct 70B
Meta 로고Meta
8
70B
128k
US$0.7
90
FireworksGoogleGroq
+10
Llama 3.1 Instruct 405B
Meta 로고Meta
7
405B
128k
-
-
-
Llama 3.1 Instruct 8B
Meta 로고Meta
7
8B
128k
US$0.0
136
GroqNovitaCoreWeave
+3
Llama 3.1 Instruct 70B
Meta 로고Meta
7
70B
128k
US$0.6
52
Amazon BedrockDeepInfra
Llama 3.2 Instruct 90B (Vision)
Meta 로고Meta
6
90B
128k
-
-
-
Llama 2 Chat 7B
Meta 로고Meta
6
7B
4k
US$0.1
-
Replicate
Llama 3.2 Instruct 3B
Meta 로고Meta
6
3B
128k
-
-
-
Llama 3 Instruct 70B
Meta 로고Meta
5
70B
8k
US$0.9
-
NovitaAmazon BedrockReplicate
Llama 3.2 Instruct 11B (Vision)
Meta 로고Meta
5
11B
128k
US$0.3
17
DeepInfra
Llama 2 Chat 70B
Meta 로고Meta
5
70B
4k
-
-
-
Llama 2 Chat 13B
Meta 로고Meta
5
13B
4k
-
-
-
Llama 65B
Meta 로고Meta
5
65B
2k
-
-
-
Llama 3.2 Instruct 1B
Meta 로고Meta
5
1B
128k
-
-
-
Llama 3 Instruct 8B
Meta 로고Meta
5
8B
8k
US$0.1
-
NovitaAmazon BedrockReplicateDeepInfra