Qwen3.8 27B vs. Mistral Small 4:リリース比較
Qwen3.8 27BとMistral Small 4のリリース比較:それぞれ4個と2個のモデルがあり、知能・性能・料金の特性が異なります。以下では、全6個のモデルの主要指標を比較します。
- 知能が最も高いのはQwen3.8 27Bです。Qwen3.8 27B (xhigh)(34)に対し、Mistral Small 4はMistral Small 4 (Reasoning)(11)です。
- 出力速度が最も速いのはMistral Small 4です。Mistral Small 4 (Reasoning)(166トークン/秒)に対し、Qwen3.8 27BはQwen3.8 27B (low)(51トークン/秒)です。
- タスクあたりの費用が最も低いのはMistral Small 4です。Mistral Small 4 (Reasoning)($0.05)に対し、Qwen3.8 27BはQwen3.8 27B (xhigh)($0.82)です。
比較するリリース
知能
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
費用
Cost per Intelligence Index Task
速度と遅延
Output Speed
能力スコア
能力指数
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
詳細
入力モダリティ | 出力モダリティ | ウェイト | プロバイダーのベンチマーク | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Qwen3.8 27B (xhigh) | 34 | 34 | 37 | 35 | 34 | 33 | 43 | $0.82 | $0.4 | $0.50 | $3.00 | $0.05 | $1,170 | 67k | 48k | 198M | 41 | 4.00秒 | 52.49秒 | 64.61秒 | 1340.79秒 | 27B | 256k | 2026年8月 | - | はい | 対応:テキスト、画像、動画 | 対応:テキスト | +2 | ||||
| Qwen3.8 27B (medium) | 28 | 28 | 34 | 27 | - | 25 | 31 | $0.90 | $0.4 | $0.50 | $3.00 | $0.05 | $977 | 52k | 32k | 113M | 49 | 3.86秒 | 44.49秒 | 54.65秒 | 999.15秒 | 27B | 256k | 2026年8月 | - | はい | 対応:テキスト、画像、動画 | 対応:テキスト | |||||
| Qwen3.8 27B (low) | 26 | 27 | 29 | 28 | - | 24 | 32 | $0.83 | $0.4 | $0.50 | $3.00 | $0.05 | $876 | 45k | 26k | 82M | 51 | 3.84秒 | 43.06秒 | 52.86秒 | 834.86秒 | 27B | 256k | 2026年8月 | - | はい | 対応:テキスト、画像、動画 | 対応:テキスト | |||||
| Qwen3.8 27B (Non-reasoning) | 20 | 19 | 16 | 22 | - | 22 | 30 | $1.91 | $0.4 | $0.50 | $3.00 | $0.05 | $1,570 | 30k | 0 | 43M | 49 | 3.87秒 | 3.87秒 | 14.06秒 | 560.61秒 | 27B | 256k | 2026年8月 | - | いいえ | 対応:テキスト、画像、動画 | 対応:テキスト | |||||
| Mistral Small 4 (Reasoning) | 11 | 10 | 7 | 13 | - | 16 | 18 | $0.05 | $0.2 | $0.15 | $0.60 | - | $79 | 14k | 9k | 54M | 166 | 0.73秒 | 12.80秒 | 15.82秒 | 93.03秒 | 119B 推論時に6.5Bが有効 | 256k | 2026年3月 | - | はい | 対応:テキスト、画像 | 対応:テキスト | |||||
| Mistral Small 4 (Non-reasoning) | 9 | - | - | - | - | - | - | - | $0.2 | $0.15 | $0.60 | - | - | - | - | - | 151 | 0.73秒 | 0.73秒 | 4.04秒 | - | 119B 推論時に6.5Bが有効 | 256k | 2026年3月 | - | いいえ | 対応:テキスト、画像 | 対応:テキスト | |||||