中型开放权重 AI 模型比较(40B-150B)

参数量介于 40B 和 150B 之间的开放权重 AI 模型。

如果模型权重可供下载,我们便将其视为开放权重模型。这样用户就可以在自己的基础设施上自行托管,并通过微调等方式定制模型。

如需了解包括方法论在内的更多详情,请参阅常见问题。

InclusionAI 标志Ling-3.0-flash-VL 和 InclusionAI 标志Ling-3.0-flash-Fin 是智能得分最高的中型开放权重模型,定义为参数量介于 40B 和 150B 之间的模型,其次是 InclusionAI 标志Ling 3.0 Flash 和 Alibaba 标志Qwen3.5 122B A10B (non-reasoning)。
Artificial Analysis Openness Index · Higher is better
Artificial Analysis Intelligence Index · Higher is better
可训练参数(十亿)

开放性

Artificial Analysis 开放性指数:得分

开放性指数以 0 至 100 的标准化评分衡量模型的开放程度(分数越高,越开放)

智能

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
查看更多

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

SciCodeUnder review

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, Criterion Pass Rate

Agentic business operations

Agentic scientific research workflows in a terminal

No data available

Quantitative analysis on spreadsheets & documents

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

规模

模型规模:总参数量与活跃参数量

Comparison between total model parameters and parameters active during inference (billions)

Intelligence Index 与活跃参数

Artificial Analysis Intelligence Index · Active parameters at inference time (billions)
Most attractive quadrant
Pareto line

Intelligence Index 与总参数量

Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line

上下文窗口

上下文窗口

Context window: tokens limit · Higher is better

更多详情

权重
服务商基准测试
Ling-3.0-flash-VL
InclusionAI 标志InclusionAI
25
124B
推理时启用 5.5B 个参数
262k
$0.0
145
InclusionAI
Ling-3.0-flash-Fin
InclusionAI 标志InclusionAI
23
124B
推理时启用 5.1B 个参数
262k
$0.0
322
InclusionAI
Qwen3.5 122B A10B (Non-reasoning)
Alibaba 标志Alibaba
18
125B
推理时启用 10B 个参数
262k
$0.7
143
DeepInfraAlibaba Cloud
Qwen3.5 122B A10B (Reasoning)
Alibaba 标志Alibaba
16
125B
推理时启用 10B 个参数
262k
$0.7
130
SiliconFlowDeepInfraAlibaba Cloud
+2
Mistral Medium 3.5
Mistral 标志Mistral
14
128B
256k
$1.2
165
MistralSelf-hosted
Nemotron 3 Super 120B A12B (Reasoning)
NVIDIA 标志NVIDIA
13
120.6B
推理时启用 12.7B 个参数
1M
$0.3
151
CrusoeNebiusDeepInfra
HyperNova 60B 2605 (High, Based on gpt-oss-120b)
Multiverse Computing 标志Multiverse Computing
12
58.7B
推理时启用 4.8B 个参数
131k
-
-
-
gpt-oss-120b (High)
OpenAI 标志OpenAI
12
117B
推理时启用 5.1B 个参数
131k
$0.2
179
GroqCloudflareScaleway
+17
K2 Think V2
Institute of Foundation Models 标志Institute of Foundation Models
11
70B
262k
-
-
-
LongCat Flash Lite
LongCat 标志LongCat
11
68.5B
推理时启用 3B 个参数
256k
-
-
-
Mistral Small 4 (Reasoning)
Mistral 标志Mistral
11
119B
推理时启用 6.5B 个参数
256k
$0.1
162
Mistral
Qwen3 Next 80B A3B (Reasoning)
Alibaba 标志Alibaba
11
80B
推理时启用 3B 个参数
262k
$0.3
177
GMIGoogleAlibaba CloudHyperbolic