日语——AI 模型基准测试 比较多语言 LLM 表现
适用于日语任务的前 5 个 AI 模型是 Gemini 3.1 Pro Preview、Gemini 3 Pro Preview (high)、Claude Opus 4.6 (max)、Claude Opus 4.6 (high) 和 Claude Opus 4.5。它们在 Artificial Analysis 多语言指数中取得了最高的日语推理得分。
如需比较所有支持语言的表现,请查看完整的多语言 AI 模型基准测试页面。
🇯🇵 适用于日语的顶尖模型
#1
Gemini 3.1 Pro Preview94#2
Gemini 3 Pro Preview (high)93#3
Claude Opus 4.6 (max)93#4
Claude Opus 4.6 (high)93#5
Claude Opus 4.593
亮点
多语言指数
多语言指数:日语
Artificial Analysis Multilingual Index · Higher is better
Reasoning models are indicated by a lightbulb icon
多语言指数:日语与价格
Artificial Analysis Multilingual Index · USD per 1M tokens (blended)
Most attractive quadrant
Reasoning models are indicated by a lightbulb icon
多语言指数:日语与输出速度
Artificial Analysis Multilingual Index · Output speed: output tokens per second
Most attractive quadrant
Reasoning models are indicated by a lightbulb icon
多语言指数:日语与上下文窗口
Artificial Analysis Multilingual Index · Context window: tokens limit
Most attractive quadrant
Reasoning models are indicated by a lightbulb icon
Global-MMLU-Lite
多语言 Global-MMLU-Lite:日语
Multilingual Global-MMLU-Lite · Higher is better
Reasoning models are indicated by a lightbulb icon
价格
Pricing: Cache Hit, Input, and Output
Price (USD per M Tokens)
Reasoning models are indicated by a lightbulb icon
速度与延迟
Output Speed
Output tokens per second · Higher is better
Reasoning models are indicated by a lightbulb icon
Latency: Time To First Answer Token
Seconds to first answer token received · Accounts for reasoning model 'thinking' time
Reasoning models are indicated by a lightbulb icon
End-to-End Response Time
Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better
Reasoning models are indicated by a lightbulb icon