Microsoft Azure: Intelligenz, Leistung und Preise der Modelle
Diese Analyse soll Sie bei der Auswahl des besten von Microsoft Azure angebotenen Modells für Ihren Anwendungsfall unterstützen.
Höchste Intelligenz
UpdatedIntelligence Index
Insgesamt 82 Modelle
Am schnellsten
Ausgabegeschwindigkeit
Insgesamt 82 Modelle
Niedrigster Preis
Mischpreis (pro 1 Mio. Tokens)
Insgesamt 82 Modelle
Azure bietet 82 Modelle mit jeweils unterschiedlichen Merkmalen bei Intelligenz, Leistung und Preis an. Nachfolgend werden die wichtigsten Metriken der Modelle verglichen.
- Bei der Intelligenz sind GPT-6 Astra (max) (53), Claude Fable 5 (with fallback) (50) und GPT-6 Sol (max) (48) die führenden Modelle von Azure.
- Bei der Ausgabegeschwindigkeit sind Llama 4 Maverick (FP8) (436 t/s), gpt-oss-120b (low) (335 t/s) und GPT-5.4 mini (xhigh) (328 t/s) am schnellsten. Die Geschwindigkeit unterscheidet sich deutlich zwischen den Modellen; der Unterschied zwischen dem schnellsten und dem langsamsten beträgt 52 %.
- Bei der Latenz bieten Phi-4 Mini (0.83 s), Llama 4 Scout (0.84 s) und Claude 4.5 Haiku (non-reasoning) (0.93 s) die kürzeste Zeit bis zum ersten Antworttoken.
- Beim Preis bieten GPT-5 nano (high) ($0.05), GPT-5 nano (medium) ($0.05) und GPT-6 Luna (max) ($0.08) die niedrigsten Mischpreise pro 1 Mio. Tokens. Die Preise unterscheiden sich zwischen den Modellen bis zum 2.7-Fachen.
- Bei der Größe des Kontextfensters unterstützen GPT-6 Sol (max) (1M), GPT-5.4 (xhigh) (1M) und Claude Fable 5 (with fallback) (1M) die größten Kontextfenster von Azure.
Intelligenzevaluationen
Artificial Analysis Intelligence Index
Intelligence Evaluations
Agentic knowledge work, (Elo-500)/2000
Agentic real-world work tasks, (Elo-500)/2000
Agentic SaaS workflows
Agentic coding & terminal use
Coding
Reasoning & knowledge
Professional document reasoning, All-pass
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Legal agentic work, criterion pass rate
Agentic business operations
Agentic scientific research workflows in a terminal
Quantitative analysis on spreadsheets & documents
Kubernetes incident root-cause analysis
Visual reasoning
Medical long context reasoning
Intelligence Index vs. Preis
Kontextfenster
Kontextfenster
Preise
Intelligence Index vs. Preis
Leistungsübersicht
Ausgabegeschwindigkeit vs. Preis
Geschwindigkeit
Gemessen anhand der Ausgabegeschwindigkeit (Tokens pro Sekunde)
Ausgabegeschwindigkeit
Latenz
Gemessen anhand der Zeit (Sekunden) bis zum ersten Token
Latenz: Zeit bis zum ersten Antworttoken
Ende-zu-Ende-Antwortzeit
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
Ende-zu-Ende-Antwortzeit vs. Preis
Weitere Analyse | |||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|
GPT-6 Astra (max) | 524k | Proprietär | 53 | $3.19 | 69 | 354.77 | 362.00 | -- | |||
Claude Fable 5 (with fallback) | 1M | Proprietär | 50 | $8.50 | 66 | 100.61 | 108.24 | -- | |||
GPT-6 Sol (max) | 1.05M | Proprietär | 48 | -- | 149 | 99.81 | 103.16 | -- | |||
GPT-5.6 Sol (max) | 524k | Proprietär | 47 | $1.95 | 103 | 127.37 | 132.24 | -- | |||
GPT-6 Astra (low) | 524k | Proprietär | 46 | $0.79 | 56 | 3.10 | 12.11 | -- | |||
GPT-5.6 Sol (xhigh) | 524k | Proprietär | 44 | $1.26 | 98 | 66.87 | 71.95 | -- | |||
GPT-5.6 Sol (high) | 524k | Proprietär | 42 | $0.85 | 88 | 17.16 | 22.86 | -- | |||
Claude Opus 4.7 (max) | 1M | Proprietär | 41* | -- | 46 | 20.67 | 31.46 | -- | |||
GPT-5.4 (xhigh) | 1.05M | Proprietär | 39* | -- | 100 | 181.33 | 186.34 | -- | |||
Claude Sonnet 5 (max) | 1M | Proprietär | 38 | $4.75 | 85 | 194.75 | 200.63 | -- | |||
GPT-6 Luna (max) | 524k | Proprietär | 38 | -- | 206 | 70.92 | 73.35 | -- | |||
GPT-5.6 Terra (xhigh) | 524k | Proprietär | 38 | $0.67 | 116 | 45.63 | 49.93 | -- | |||
GPT-5.5 (high) | 524k | Proprietär | 37 | $1.13 | 83 | 35.33 | 41.33 | -- | |||
GPT-5.6 Luna (xhigh) | 524k | Proprietär | 35 | $0.09 | 162 | 46.38 | 49.47 | -- | |||
GPT-5.6 Terra (high) | 524k | Proprietär | 34 | $0.35 | 109 | 3.09 | 7.67 | -- | |||
GPT-5.6 Luna (high) | 524k | Proprietär | 32 | $0.05 | 150 | 13.55 | 16.89 | -- | |||
Claude Opus 4.6 (max) | 1M | Proprietär | 32* | -- | 41 | 24.77 | 37.10 | -- | |||
DeepSeek V4 Pro (max) | 1M | Offen | 30 | $9.65 | 81 | 1.72 | 62.11 | 54.20 | |||
GPT-5.2 (xhigh) | 400k | Proprietär | 30* | -- | 89 | 110.20 | 115.85 | -- | |||
DeepSeek V4 Pro (high) | 1M | Offen | 30* | -- | 82 | 1.85 | 32.08 | 24.17 | |||
Claude Sonnet 4.6 (max) | 200k | Proprietär | 30 | $2.70 | 56 | 130.59 | 139.48 | -- | |||
Claude Opus 4.5 | 200k | Proprietär | 29* | -- | 48 | 18.72 | 29.21 | -- | |||
GPT-5.2 Codex (xhigh) | 400k | Proprietär | 29* | -- | 115 | 86.52 | 90.87 | -- | |||
Kimi K2.6 | 262k | Offen | 27 | $1.91 | 227 | 1.53 | 23.34 | 19.61 | |||
GPT-5.2 (medium) | 400k | Proprietär | 27* | -- | 88 | 5.18 | 10.87 | -- | |||
Claude Opus 4.6 (non-reasoning, high) | 1M | Proprietär | 26* | -- | 38 | 1.88 | 15.12 | -- | |||
Grok 4.20 0309 v2 | 262k | Proprietär | 26* | -- | 216 | 10.56 | 12.87 | -- | |||
GPT-5 Codex (high) | 400k | Proprietär | 25* | -- | 156 | 10.45 | 13.65 | -- | |||
Grok 4.3 (high) | 200k | Proprietär | 25 | $0.53 | 157 | 20.96 | 24.14 | -- | |||
Grok 4.3 (medium) | 200k | Proprietär | 25* | -- | 157 | 14.72 | 17.91 | -- | |||
GPT-5.1 (high) | 272k | Proprietär | 25* | -- | 115 | 41.64 | 45.99 | -- | |||
Claude Sonnet 4.6 (non-reasoning, high) | 200k | Proprietär | 25* | -- | 43 | 1.63 | 13.15 | -- | |||
Grok 4.3 (low) | 200k | Proprietär | 24* | -- | 139 | 6.02 | 9.61 | -- | |||
GPT-5.4 mini (xhigh) | 400k | Proprietär | 24 | $0.43 | 254 | 127.11 | 129.08 | -- | |||
GPT-5.1 Codex (high) | 400k | Proprietär | 24* | -- | 120 | 9.35 | 13.50 | -- | |||
Claude Opus 4.5 (non-reasoning) | 200k | Proprietär | 24* | -- | 47 | 1.51 | 12.13 | -- | |||
Kimi K2.6 (non-reasoning) | 262k | Offen | 24* | -- | 194 | 1.60 | 4.18 | -- | |||
Kimi K2.5 | 262k | Offen | 23* | -- | 131 | 1.40 | 27.93 | 22.70 | |||
GPT-5 (high) | 400k | Proprietär | 23* | -- | 97 | 68.03 | 73.21 | -- | |||
GPT-5 (medium) | 400k | Proprietär | 23* | -- | 91 | 35.09 | 40.55 | -- | |||
Claude 4.1 Opus | 200k | Proprietär | 23* | -- | -- | -- | -- | -- | |||
Grok 4 | 256k | Proprietär | 22* | -- | 71 | 10.93 | 17.93 | -- | |||
Kimi K2 Thinking | 256k | Offen | 22* | -- | 99 | 1.91 | 27.12 | 20.17 | |||
o3-pro | 200k | Proprietär | 22* | -- | -- | -- | -- | -- | |||
DeepSeek V4 Pro (non-reasoning) | 1M | Offen | 21* | -- | 89 | 1.58 | 7.21 | -- | |||
GPT-5 (low) | 400k | Proprietär | 21* | -- | 92 | 11.44 | 16.85 | -- | |||
Claude 4.5 Sonnet | 200k | Proprietär | 21 | $0.54 | 44 | 14.87 | 26.12 | -- | |||
GPT-5 mini (medium) | 400k | Proprietär | 21* | -- | 132 | 10.36 | 14.15 | -- | |||
GPT-5.1 Codex mini (high) | 400k | Proprietär | 20* | -- | 163 | 15.60 | 18.67 | -- | |||
o3 | 200k | Proprietär | 20* | -- | 97 | 24.29 | 29.43 | -- | |||
GPT-5.4 mini (medium) | 400k | Proprietär | 20* | -- | 211 | 9.14 | 11.51 | -- | |||
Kimi K2.5 (non-reasoning) | 262k | Offen | 19* | -- | 122 | 1.35 | 5.46 | -- | |||
Claude 4.5 Sonnet (non-reasoning) | 200k | Proprietär | 19* | -- | 42 | 1.46 | 13.31 | -- | |||
Claude 4.1 Opus (non-reasoning) | 200k | Proprietär | 19* | -- | -- | -- | -- | -- | |||
Grok 4 Fast | 2M | Proprietär | 18* | -- | -- | -- | -- | -- | |||
GPT-5.2 (non-reasoning) | 400k | Proprietär | 17* | -- | 86 | 1.48 | 7.31 | -- | |||
Claude 4.5 Haiku | 200k | Proprietär | 17 | $0.22 | 98 | 15.36 | 20.47 | -- | |||
GPT-5 mini (high) | 400k | Proprietär | 17 | $0.05 | 131 | 53.20 | 57.02 | -- | |||
o4-mini (high) | 200k | Proprietär | 17* | -- | 138 | 22.09 | 25.71 | -- | |||
DeepSeek V3.2 (non-reasoning) | 128k | Offen | 16* | -- | 166 | 1.90 | 4.91 | -- | |||
Claude 4.5 Haiku (non-reasoning) | 200k | Proprietär | 15* | -- | 90 | 1.11 | 6.68 | -- | |||
o1 | 200k | Proprietär | 15* | -- | -- | -- | -- | -- | |||
Grok 3 mini Reasoning (high) | 32k | Proprietär | 15* | -- | -- | -- | -- | -- | |||
GPT-5.1 (non-reasoning) | 400k | Proprietär | 13* | -- | 110 | 1.24 | 5.80 | -- | |||
GPT-5 nano (high) | 400k | Proprietär | 13* | -- | 203 | 69.34 | 71.80 | -- | |||
GPT-4.1 | 1M | Proprietär | 13* | -- | 114 | 1.64 | 6.02 | -- | |||
GPT-5 nano (medium) | 400k | Proprietär | 12* | -- | 205 | 32.76 | 35.20 | -- | |||
o3-mini | 200k | Proprietär | 12* | -- | 218 | 7.00 | 9.29 | -- | |||
Grok 3 | 16k | Proprietär | 12* | -- | -- | -- | -- | -- | |||
gpt-oss-120b (high) | 131k | Offen | 12 | $0.11 | 315 | 0.72 | 8.65 | 6.34 | |||
GPT-5 (minimal) | 400k | Proprietär | 11* | -- | 93 | 1.78 | 7.14 | -- | |||
o1-preview | 128k | Proprietär | 11* | -- | -- | -- | -- | -- | |||
GPT-5.4 mini (non-reasoning) | 400k | Proprietär | 11* | -- | 187 | 1.05 | 3.73 | -- | |||
Grok 4 Fast (non-reasoning) | 2M | Proprietär | 11* | -- | -- | -- | -- | -- | |||
o3-mini (high) | 200k | Proprietär | 11 | -- | 228 | 20.21 | 22.40 | -- | |||
gpt-oss-120b (low) | 131k | Offen | 10* | -- | 334 | 0.84 | 8.32 | 5.99 | |||
GPT-4.1 mini | 1M | Proprietär | 10* | -- | 103 | 1.54 | 6.38 | -- | |||
Llama 4 Maverick (FP8) | 128k | Offen | 10* | -- | 455 | 1.28 | 2.38 | -- | |||
GPT-5 mini (minimal) East US 2 - Global Standard | 400k | Proprietär | 10* | -- | -- | -- | -- | -- | |||
Mistral Large 3 | 256k | Offen | 9 | $0.10 | 56 | 1.86 | 10.78 | -- | |||
Mistral Medium 3 | 128k | Proprietär | 9* | -- | 47 | 2.27 | 12.91 | -- | |||
GPT-4o (Nov) | 128k | Proprietär | 8* | -- | 125 | 1.90 | 5.91 | -- | |||
Llama 4 Scout | 128k | Offen | 8* | -- | 134 | 0.85 | 4.57 | -- | |||
GPT-4.1 nano | 1M | Proprietär | 8* | -- | 276 | 1.47 | 3.27 | -- | |||
GPT-4o (Aug) | 128k | Proprietär | 8* | -- | 141 | 1.38 | 4.93 | -- | |||
Llama 3.3 70B | 128k | Offen | 8* | -- | 124 | 2.18 | 6.20 | -- | |||
GPT-4o (May) | 128k | Proprietär | 7* | -- | 143 | 1.55 | 5.03 | -- | |||
GPT-5 nano (minimal) East US 2 - Global Standard | 400k | Proprietär | 7* | -- | -- | -- | -- | -- | |||
GPT-4 Turbo | 128k | Proprietär | 7* | -- | 108 | 1.78 | 6.42 | -- | |||
Command A | 256k | Offen | 7* | -- | 42 | 3.04 | 14.82 | -- | |||
GPT-4o mini | 128k | Proprietär | 7* | -- | 78 | 1.87 | 8.24 | -- | |||
Phi-4 Mini | 128k | Offen | 6* | -- | 44 | 0.82 | 12.09 | -- | |||
Phi-4 | 16.4k | Offen | 6* | -- | 36 | 2.79 | 16.59 | -- | |||
Phi-4 Multimodal | 4.1k | Offen | 6* | -- | -- | -- | -- | -- | |||
Wichtige Definitionen
Häufig gestellte Fragen
Häufige Fragen zu Microsoft Azure
Microsoft Azure bietet 82 von uns erfasste Modelle an: GPT-6 Astra (max), Claude Fable 5 (with fallback), GPT-6 Sol (max), GPT-5.6 Sol (max), GPT-6 Astra (low), GPT-5.6 Sol (xhigh), GPT-5.6 Sol (high), Claude Opus 4.7 (max), GPT-5.4 (xhigh), Claude Sonnet 5 (max), GPT-6 Luna (max), GPT-5.6 Terra (xhigh), GPT-5.5 (high), GPT-5.6 Luna (xhigh), GPT-5.6 Terra (high), GPT-5.6 Luna (high), Claude Opus 4.6 (max), DeepSeek V4 Pro (max), GPT-5.2 (xhigh), DeepSeek V4 Pro (high), Claude Sonnet 4.6 (max), Claude Opus 4.5, GPT-5.2 Codex (xhigh), Kimi K2.6, GPT-5.2 (medium), Claude Opus 4.6 (non-reasoning, high), Grok 4.20 0309 v2, GPT-5 Codex (high), Grok 4.3 (high), Grok 4.3 (medium), GPT-5.1 (high), Claude Sonnet 4.6 (non-reasoning, high), Grok 4.3 (low), GPT-5.4 mini (xhigh), GPT-5.1 Codex (high), Claude Opus 4.5 (non-reasoning), Kimi K2.6 (non-reasoning), Kimi K2.5, GPT-5 (high), GPT-5 (medium), Grok 4, Kimi K2 Thinking, DeepSeek V4 Pro (non-reasoning), GPT-5 (low), Claude 4.5 Sonnet, GPT-5 mini (medium), GPT-5.1 Codex mini (high), o3, GPT-5.4 mini (medium), Kimi K2.5 (non-reasoning), Claude 4.5 Sonnet (non-reasoning), GPT-5.2 (non-reasoning), Claude 4.5 Haiku, GPT-5 mini (high), o4-mini (high), DeepSeek V3.2 (non-reasoning), Claude 4.5 Haiku (non-reasoning), GPT-5.1 (non-reasoning), GPT-5 nano (high), GPT-4.1, GPT-5 nano (medium), o3-mini, gpt-oss-120b (high), GPT-5 (minimal), GPT-5.4 mini (non-reasoning), o3-mini (high), gpt-oss-120b (low), GPT-4.1 mini, Llama 4 Maverick (FP8), Mistral Large 3, Mistral Medium 3, GPT-4o (Nov), Llama 4 Scout, GPT-4.1 nano, GPT-4o (Aug), Llama 3.3 70B, GPT-4o (May), GPT-4 Turbo, Command A, GPT-4o mini, Phi-4 Mini und Phi-4.
Das intelligenteste bei Microsoft Azure verfügbare Modell ist GPT-6 Astra (max) mit einem Intelligence-Index-Wert von 53.
Gemessen an der Ausgabegeschwindigkeit ist Llama 4 Maverick (FP8) mit 436.5 Tokens pro Sekunde das schnellste Modell von Microsoft Azure.
Das Modell von Microsoft Azure mit der kürzesten Zeit bis zum ersten Antworttoken ist Phi-4 Mini mit 0.83 s. Eine niedrigere Latenz bedeutet eine schnellere erste Antwort.
Gemessen am Mischpreis ist GPT-5 nano (high) mit $0.05 pro 1 Mio. Tokens (Verhältnis von Cache-Treffern, Eingabe und Ausgabe: 7:2:1) das günstigste Modell von Microsoft Azure.
Die Preise der Modelle von Microsoft Azure unterscheiden sich bis zum 224-Fachen: von $0.05 pro 1 Mio. Tokens für GPT-5 nano (high) bis $12.00 pro 1 Mio. Tokens für GPT-4 Turbo.
Ja, Microsoft Azure bietet eine OpenAI-kompatible API. Das erleichtert den Wechsel von OpenAI und die Verwendung bestehender Integrationen mit dem OpenAI SDK.
69 von 82 Modellen von Microsoft Azure unterstützen den JSON-Modus für strukturierte Ausgaben.
80 von 82 Modellen von Microsoft Azure unterstützen Funktionsaufrufe (Tool-Nutzung).
Ja, Microsoft Azure bietet 53 Reasoning-Modelle an: GPT-6 Astra (max), Claude Fable 5 (with fallback), GPT-6 Sol (max), GPT-5.6 Sol (max), GPT-6 Astra (low), GPT-5.6 Sol (xhigh), GPT-5.6 Sol (high), Claude Opus 4.7 (max), GPT-5.4 (xhigh), Claude Sonnet 5 (max), GPT-6 Luna (max), GPT-5.6 Terra (xhigh), GPT-5.5 (high), GPT-5.6 Luna (xhigh), GPT-5.6 Terra (high), GPT-5.6 Luna (high), Claude Opus 4.6 (max), DeepSeek V4 Pro (max), GPT-5.2 (xhigh), DeepSeek V4 Pro (high), Claude Sonnet 4.6 (max), Claude Opus 4.5, GPT-5.2 Codex (xhigh), Kimi K2.6, GPT-5.2 (medium), Grok 4.20 0309 v2, GPT-5 Codex (high), Grok 4.3 (high), Grok 4.3 (medium), GPT-5.1 (high), Grok 4.3 (low), GPT-5.4 mini (xhigh), GPT-5.1 Codex (high), Kimi K2.5, GPT-5 (high), GPT-5 (medium), Grok 4, Kimi K2 Thinking, GPT-5 (low), Claude 4.5 Sonnet, GPT-5 mini (medium), GPT-5.1 Codex mini (high), o3, GPT-5.4 mini (medium), Claude 4.5 Haiku, GPT-5 mini (high), o4-mini (high), GPT-5 nano (high), GPT-5 nano (medium), o3-mini, gpt-oss-120b (high), o3-mini (high) und gpt-oss-120b (low). Reasoning-Modelle nutzen ausführliches Schlussfolgern, um komplexe Probleme vor der Antwort zu bearbeiten.
Ja, 18 von 82 Modellen von Microsoft Azure haben offene Gewichte: DeepSeek V4 Pro (max), DeepSeek V4 Pro (high), Kimi K2.6, Kimi K2.6 (non-reasoning), Kimi K2.5, Kimi K2 Thinking, DeepSeek V4 Pro (non-reasoning), Kimi K2.5 (non-reasoning), DeepSeek V3.2 (non-reasoning), gpt-oss-120b (high), gpt-oss-120b (low), Llama 4 Maverick (FP8), Mistral Large 3, Llama 4 Scout, Llama 3.3 70B, Command A, Phi-4 Mini und Phi-4.
Ja, die Leistung eines Anbieters kann sich aufgrund von Infrastrukturänderungen, Lastverteilung und Aktualisierungen im Laufe der Zeit verändern. Wir benchmarken alle Anbieter fortlaufend und zeigen historische Leistungstrends in den Diagrammen „Im Zeitverlauf“ an.
Berücksichtigen Sie bei der Auswahl eines Modells von Microsoft Azure die Intelligenz für qualitätssensible Aufgaben, die Ausgabegeschwindigkeit für durchsatzintensive Aufgaben, die Latenz für interaktive Anwendungen mit schnellen ersten Antworten, die Preise für kostensensible Workloads sowie Funktionen wie die Größe des Kontextfensters, den JSON-Modus und die Unterstützung von Funktionsaufrufen.