Parasail: Intelligenz, Leistung und Preise der Modelle

Parasail
Parasail

Diese Analyse soll Sie bei der Auswahl des besten von Parasail angebotenen Modells für Ihren Anwendungsfall unterstützen.

Höchste Intelligenz

Updated
#1
GLM-5.3 (max)GLM-5.3 (max)
45
#2
Kimi K3 (max)Kimi K3 (max)
44
#3
GLM-5.3-FlashGLM-5.3-Flash
42
#4
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
35
#5
GLM-5.2 (max) (NVFP4)GLM-5.2 (max) (NVFP4)
34

Intelligence Index

Insgesamt 27 Modelle

Am schnellsten

#1
GLM-5.2 (max) (NVFP4)GLM-5.2 (max) (NVFP4)
187 t/s
#2
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
187 t/s
#3
Kimi K2.6 (Non-reasoning) (INT4)Kimi K2.6 (Non-reasoning) (INT4)
173 t/s
#4
Qwen3 Next 80B A3BQwen3 Next 80B A3B
173 t/s
#5
Kimi K2.6Kimi K2.6
169 t/s

Ausgabegeschwindigkeit

Insgesamt 27 Modelle

Niedrigster Preis

#1
Gemma 3 27BGemma 3 27B
$0.09
#2
GLM-5.3-FlashGLM-5.3-Flash
$0.10
#3
Gemma 4 26B A4BGemma 4 26B A4B
$0.10
#4
Gemma 4 26B A4B (Non-reasoning)Gemma 4 26B A4B (Non-reasoning)
$0.10
#5
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
$0.11

Mischpreis (pro 1 Mio. Tokens)

Insgesamt 27 Modelle

Parasail bietet 27 Modelle mit jeweils unterschiedlichen Merkmalen bei Intelligenz, Leistung und Preis an. Nachfolgend werden die wichtigsten Metriken der Modelle verglichen.

  • Bei der Intelligenz sind GLM-5.3 (max) (45), Kimi K3 (max) (44) und GLM-5.3-Flash (42) die führenden Modelle von Parasail.
  • Bei der Ausgabegeschwindigkeit sind GLM-5.2 (max) (NVFP4) (187 t/s), DeepSeek V4 Flash 0731 (max) (187 t/s) und Kimi K2.6 (Non-reasoning) (INT4) (173 t/s) am schnellsten.
  • Bei der Latenz bieten Qwen3.6 35B A3B (Non-reasoning) (FP8) (0.87 s), Llama 4 Maverick (FP8) (0.95 s) und Qwen3 Coder Next (FP8) (1.03 s) die kürzeste Zeit bis zum ersten Antworttoken.
  • Beim Preis bieten Gemma 3 27B ($0.09), GLM-5.3-Flash ($0.10) und Gemma 4 26B A4B ($0.10) die niedrigsten Mischpreise pro 1 Mio. Tokens.
  • Bei der Größe des Kontextfensters unterstützen Kimi K3 (max) (1M), DeepSeek V4 Flash 0731 (max) (1M) und DeepSeek V4 Flash (high) (FP8) (1M) die größten Kontextfenster von Parasail.

Wichtigste Ergebnisse

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Intelligenzevaluationen

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

Instruction following

Agentic tool use

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Kontextfenster

Context Window

Context window: tokens limit · Higher is better

Preise

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Leistungsübersicht

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Geschwindigkeit

Gemessen anhand der Ausgabegeschwindigkeit (Tokens pro Sekunde)

Output Speed

Output tokens per second · Higher is better

Latenz

Gemessen anhand der Zeit (Sekunden) bis zum ersten Token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

Ende-zu-Ende-Antwortzeit

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Weitere Analyse
Logo von Z AI
GLM-5.3 (max)
1M
Offen
45
$1.97
114
1.13
23.06
17.54
Logo von Kimi
Kimi K3 (max)
1.05M
Offen
44
$2.42
48
1.21
53.16
41.56
Logo von Z AI
GLM-5.3-Flash
1M
Offen
42
$0.28
107
1.22
24.54
18.65
Logo von DeepSeek
DeepSeek V4.1 Flash (max)
1.05M
Offen
40
$0.90
--
--
--
--
Logo von DeepSeek
DeepSeek V4 Flash 0731 (max)
1.05M
Offen
35
$0.41
187
1.48
14.85
10.70
Logo von Z AI
GLM-5.2 (max) (NVFP4)
1M
Offen
34
--
187
1.14
14.50
10.69
Logo von Alibaba
Qwen3.8 27B (xhigh) (FP8)
262k
Offen
34
$2.29
84
1.32
31.23
23.93
Logo von Kimi
Kimi K2.6
262k
Offen
31*
--
169
1.18
30.46
26.32
Logo von MiniMax
MiniMax-M3 (MXFP8)
1M
Offen
30
$0.46
129
0.93
20.25
15.45
Logo von DeepSeek
DeepSeek V4 Flash (high) (FP8)
1.05M
Offen
25
--
70
1.35
26.17
17.69
Logo von DeepSeek
DeepSeek V4 Flash (max) (FP8)
1.05M
Offen
25
--
74
1.36
84.00
75.88
Logo von Kimi
Kimi K2.6 (Non-reasoning) (INT4)
262k
Offen
24*
--
173
1.20
4.09
--
Logo von Alibaba
Qwen3.5 397B A17B
262k
Offen
19
$0.30
52
0.96
71.67
61.12
Logo von Alibaba
Qwen3.6 35B A3B
262k
Offen
19
$0.12
124
0.82
48.21
43.38
Logo von Google
Gemma 4 26B A4B
256k
Offen
17*
--
23
1.85
109.65
86.24
Logo von Google
Gemma 4 31B
262k
Offen
15
$0.05
27
2.10
84.17
63.72
Logo von Alibaba
Qwen3.6 35B A3B (Non-reasoning) (FP8)
262k
Offen
15*
--
119
0.87
5.06
--
Logo von Google
Gemma 4 31B (Non-reasoning)
262k
Offen
14*
--
33
1.42
16.39
--
Logo von Google
Gemma 4 26B A4B (Non-reasoning)
262k
Offen
13*
--
19
1.57
27.76
--
Logo von OpenAI
gpt-oss-120b (high)
131k
Offen
12
$0.06
169
0.75
15.55
11.84
Logo von Alibaba
Qwen3 235B 2507 (Non-reasoning)
262k
Offen
12*
--
32
1.15
16.90
--
Logo von OpenAI
gpt-oss-120b (low)
131k
Offen
10*
--
159
0.80
16.48
12.54
Logo von Alibaba
Qwen3 Coder Next (FP8)
262k
Offen
10
$0.12
73
1.03
7.90
--
Logo von Alibaba
Qwen3 VL 235B A22B (FP8)
131k
Offen
10*
--
45
1.18
12.21
--
Logo von Alibaba
Qwen3 Next 80B A3B
262k
Offen
10*
--
173
1.53
4.42
--
Logo von Meta
Llama 4 Maverick (FP8)
1.05M
Offen
9
$0.03
56
0.95
9.92
--
Logo von Meta
Llama 3.3 70B (FP8)
131k
Offen
8*
--
86
2.64
8.48
--
Logo von Allen Institute for AI
Olmo 3.1 32B Think
65.5k
Offen
7*
--
--
--
--
--
Logo von Allen Institute for AI
Olmo 3 7B
65.5k
Offen
5*
--
--
--
--
--
Logo von Google
Gemma 3 27B
131k
Offen
5
$0.10
42
2.12
14.04
--

Wichtige Definitionen

Häufig gestellte Fragen

Häufige Fragen zu Parasail