GPT-6.1 Sol vs. Mistral Small 4: 릴리스 비교
GPT-6.1 Sol와 Mistral Small 4 릴리스 비교: 각각 5개, 2개 모델이며 지능, 성능, 가격 특성이 서로 다릅니다. 아래에서 전체 7개 모델의 주요 지표를 비교합니다.
- 지능 부문에서 GPT-6.1 Sol가 가장 높습니다. GPT-6.1 Sol (max)(52), Mistral Small 4는 Mistral Small 4 (Reasoning)(11)입니다.
- 출력 속도가 가장 빠른 것은 Mistral Small 4입니다. Mistral Small 4 (Reasoning)(179 토큰/초), GPT-6.1 Sol는 GPT-6.1 Sol (low)(74 토큰/초)입니다.
- 작업당 비용이 가장 낮은 것은 Mistral Small 4입니다. Mistral Small 4 (Reasoning)($0.02), GPT-6.1 Sol는 GPT-6.1 Sol (low)($0.13)입니다.
지능
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
비용
Cost per Intelligence Index Task
속도 및 지연 시간
Output Speed
역량 점수
역량 지수
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
상세 정보
입력 모달리티 | 출력 모달리티 | 웨이트 | 제공업체 벤치마크 | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| GPT-6.1 Sol (max) | 52 | 54 | 57 | 58 | - | 54 | 59 | $0.72 | US$1.5 | $2.00 | $10.00 | $0.10 | US$1,082 | 38k | 25k | 67M | 67 | 267.64초 | 267.64초 | 275.12초 | 568.66초 | - | 1M | 2026년 9월 | - | 예 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | 제공되지 않음 | - | |||
| GPT-6.1 Sol (xhigh) | 51 | 53 | 56 | 57 | - | 53 | 58 | $0.39 | US$1.5 | $2.00 | $10.00 | $0.10 | US$662 | 18k | 9k | 36M | 64 | 68.65초 | 68.65초 | 76.51초 | 271.10초 | - | 1M | 2026년 9월 | - | 예 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | 제공되지 않음 | - | |||
| GPT-6.1 Sol (high) | 50 | 53 | 55 | 56 | - | 52 | 58 | $0.32 | US$1.5 | $2.00 | $10.00 | $0.10 | US$521 | 13k | 6k | 25M | 66 | 57.26초 | 57.26초 | 64.81초 | 202.65초 | - | 1M | 2026년 9월 | - | 예 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | 제공되지 않음 | - | |||
| GPT-6.1 Sol (medium) | 48 | 50 | 53 | 54 | - | 50 | 56 | $0.21 | US$1.5 | $2.00 | $10.00 | $0.10 | US$361 | 8k | 3k | 15M | 62 | 5.29초 | 5.29초 | 13.36초 | 126.91초 | - | 1M | 2026년 9월 | - | 예 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | 제공되지 않음 | - | |||
| GPT-6.1 Sol (low) | 42 | 45 | 46 | 50 | - | 44 | 52 | $0.13 | US$1.5 | $2.00 | $10.00 | $0.10 | US$250 | 4k | 1k | 9M | 74 | 1.84초 | 1.84초 | 8.56초 | 54.07초 | - | 1M | 2026년 9월 | - | 예 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | 제공되지 않음 | - | |||
| Mistral Small 4 (Reasoning) | 11 | 10 | 7 | 13 | - | 16 | 18 | $0.02 | US$0.1 | $0.15 | $0.60 | $0.015 | US$42 | 14k | 9k | 54M | 179 | 0.78초 | 11.93초 | 14.72초 | 84.79초 | 119B 추론 시 6.5B 활성 | 256k | 2026년 3월 | - | 예 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | |||||
| Mistral Small 4 (Non-reasoning) | 9 | - | - | - | - | - | - | - | US$0.1 | $0.15 | $0.60 | $0.015 | - | - | - | - | 161 | 0.75초 | 0.75초 | 3.86초 | - | 119B 추론 시 6.5B 활성 | 256k | 2026년 3월 | - | 아니요 | 지원: 텍스트 및 이미지 | 지원: 텍스트 | |||||