Vergleich von KI-Chatbots

Konversationsagenten in Websites, Anwendungen oder Messagingkanälen, die in einer Chatoberfläche Fragen beantworten oder klar abgegrenzte Aufgaben ausführen. Meist geht es um häufig gestellte Fragen, einfache Abläufe oder Fachwissen und nicht um die langfristige eigenständige Bearbeitung von Aufgaben.

Sprachmodelle können Sie in unseren Modell-Benchmarks vergleichen.

Wir nutzen KI, um einige Ergebnisse zu erheben

Highlights

Intelligence Index · Higher is better
Feature score · Higher is better · 6 categories tracked
Monthly price (USD) · Standard & premium plans

Planoptionen vergleichen

ProduktStufePreisModellIntelligenzEingabenMedien-Gen.ToolsGedächtnisMCPKonnektorenFunktionsbewertungAppsDatenschutz
ChatGPT Plus
OpenAIOpenAI
Standard
$20/MonatGPT-5.6 Sol (max)61
BildPDFExcelSprache
BildVideoSpracheSprache-zu-Sprache
WebCodeDatenRecherche
GedächtnisVerlauf
9/104.7/6
iOSAndroidmacOSWindows
Claude Pro
AnthropicAnthropic
Standard
$20/MonatClaude Opus 5 (max)63
BildPDFExcelSprache
Sprache
WebCodeDatenRecherche
GedächtnisVerlauf
3/104.3/6
iOSAndroidmacOSWindows
Google AI Pro
GoogleGoogle
Standard
$20/MonatGemini 3.6 Flash52
BildPDFExcelVideoSprache
BildVideoSpracheSprache-zu-Sprache
WebCodeDatenRecherche
GedächtnisVerlauf
5/104.5/6
iOSAndroid
Poe Pro
PoePoe
Standard
$20/Monat
Claude Opus 4.6 (max)
45
BildPDFExcelSprache
BildVideo
Web
Verlauf
0/102.0/6
iOSAndroidmacOSWindows
Perplexity Pro
PerplexityPerplexity
Standard
$20/Monat
Gemini 3.1 Pro Preview
48
BildPDFExcelSprache
BildVideoSprache
WebCodeDatenRecherche
Verlauf
2/103.3/6
iOSAndroidmacOSWindows
SuperGrok
SpaceXAISpaceXAI
Standard
$30/MonatGrok 4.5 (high)56
BildPDFExcelSprache
BildVideoSprache
WebCodeDatenRecherche
GedächtnisVerlauf
2/103.8/6
iOSAndroid
Mistral Vibe Pro
MistralMistral
Standard
$15/MonatMistral Medium 3.530
BildPDFExcelSprache
BildSprache
WebCodeDatenRecherche
GedächtnisVerlauf
2/104.5/6
iOSAndroid

Marktüberblick

Die großen Anbieter Claude, ChatGPT, Gemini und Meta AI haben unterschiedliche Stärken beim Schlussfolgern, bei der Antwortqualität und bei Funktionen. Die meisten bieten kostenlose Tarife sowie kostenpflichtige Angebote (15–25 $ pro Monat) mit größeren Kontextfenstern, Websuche, Datei-Uploads und Bildgenerierung. Der Markt konzentriert sich auf diese führenden Anbieter; Open-Source-Optionen wie Llama und Mistral gewinnen an Bedeutung, wenn Datenschutz wichtig ist.

Unterscheidungsmerkmale

  • Einige zeichnen sich bei der Programmierung aus (CodeStral, o1), andere bei der Gesprächsgenauigkeit (Claude) oder beim visuellen Verständnis (Gemini).
  • Die Unterschiede liegen in spezialisierten Fähigkeiten wie langen Kontexten, Echtzeitsuche, Sprache und Multimodalität – nicht in der grundlegenden Gesprächsqualität, die ein vergleichbares Niveau erreicht hat.

Intelligenz

Artificial Analysis Intelligence Index: Modell mit maximaler unterstützter Intelligenz

Artificial Analysis Intelligence Index · Higher is better

Artificial Analysis Intelligence Index v4.1.1 includes: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. See Intelligence Index methodology for further details, including a breakdown of each evaluation and how we run them.

Funktionen

Intelligence Index vs. Funktionsbewertung

Feature score · Higher is better · 6 categories tracked
Most attractive quadrant

A summary score counting the features offered by chatbots across six main categories (all listed in the full comparison table above):

  • Media Generation: 1 point (image generation: 0.25, video generation: 0.25, voice conversation: 0.25, native voice-to-voice: 0.25)
  • Tools: 1 point (web search: 0.25, code interpreter: 0.25, data analysis: 0.25, deep research: 0.25)
  • Input Capabilities: 1 point (image input: 0.20, PDF input: 0.20, Excel/CSV input: 0.20, video input: 0.20, voice input: 0.20)
  • Memory: 1 point (memory: 0.5, chat history: 0.5)
  • MCP Support: 1 point (Model Context Protocol integration: 1.0)
  • Connectors: 1 point (available integrations and connectors for external services and tools; each connector weighs 0.1 points; see the 10 connectors evaluated in the comparison table above)

The maximum total value for this metric is 6 points.

Artificial Analysis Intelligence Index v4.1.1 includes: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. See Intelligence Index methodology for further details, including a breakdown of each evaluation and how we run them.

Preis

Intelligence Index vs. Preis (Bezahlpläne)

Artificial Analysis Intelligence Index · Monthly price (USD) · Standard & premium plans
Most attractive quadrant

These charts reveal the pricing structure and value proposition of different paid chatbot plans.

Key insights: The market shows a clear pricing hierarchy with three distinct tiers. Premium plans (SuperGrok Heavy, Google AI Ultra, Claude Max, Perplexity Max, ChatGPT Pro) are positioned at ~$200-300/month, offering top-tier capabilities. Standard plans cluster at much more affordable pricing ($15-30/month) with most options around $20/month, providing accessible AI capabilities for broader audiences.

Monthly Price: The monthly subscription cost in USD for standard and premium chatbot plans. Free plans are excluded from this analysis.

Artificial Analysis Intelligence Index v4.1.1 includes: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. See Intelligence Index methodology for further details, including a breakdown of each evaluation and how we run them.

Monatspreis: Bezahlte Chatbot-Pläne

Monthly price (USD) · Standard & premium plans

These charts reveal the pricing structure and value proposition of different paid chatbot plans.

Key insights: The market shows a clear pricing hierarchy with three distinct tiers. Premium plans (SuperGrok Heavy, Google AI Ultra, Claude Max, Perplexity Max, ChatGPT Pro) are positioned at ~$200-300/month, offering top-tier capabilities. Standard plans cluster at much more affordable pricing ($15-30/month) with most options around $20/month, providing accessible AI capabilities for broader audiences.

Monthly Price: The monthly subscription cost in USD for standard and premium chatbot plans. Free plans are excluded from this analysis.

Funktionsbewertung vs. Preis (Bezahlpläne)

Feature score · Higher is better · 6 categories tracked · Monthly price (USD) · Standard & premium plans
Most attractive quadrant

This scatter plot reveals the feature value proposition of different paid chatbot plans by comparing their comprehensive feature offerings against monthly pricing.

Key insights: The market shows a clear pricing hierarchy with three distinct tiers. Premium plans (SuperGrok Heavy, Google AI Ultra, Claude Max, Perplexity Max, ChatGPT Pro) are positioned at ~$200-300/month, offering top-tier capabilities. Standard plans cluster at much more affordable pricing ($15-30/month) with most options around $20/month, providing accessible AI capabilities for broader audiences. Notably, some standard plans outperform premium alternatives on features, highlighting diverse value propositions across price points.

Monthly Price: The monthly subscription cost in USD for standard and premium chatbot plans. Free plans are excluded from this analysis.

Häufig gestellte Fragen

Unser Vergleich nutzt Benchmark-Daten, darunter den Intelligence Index, Funktionsbewertungen, die Größe des Kontextfensters und den Preis. Sie können nach Modell filtern, Pläne nebeneinander vergleichen und Diagramme für Intelligenz vs. Preis sowie Intelligenz vs. Funktionen ansehen, um die beste Wahl für Ihren Anwendungsfall zu finden.

Der Intelligence Index ist ein zusammengesetzter Wert auf Basis des Benchmarkings von Artificial Analysis, der Reasoning, Wissen und Antwortqualität misst. Höhere Werte deuten auf eine stärkere Leistung bei standardisierten Bewertungen hin. Bezahlpläne schneiden in der Regel besser ab als kostenlose Stufen.

Die meisten großen Anbieter (Claude, ChatGPT, Gemini, Meta AI) bieten kostenlose Stufen mit unterschiedlichen Limits. Unsere Vergleichstabelle zeigt Funktionsbewertungen, Kontextfenster und Fähigkeiten für jeden Plan. Kostenlose Stufen haben oft kleinere Kontextfenster und weniger erweiterte Funktionen wie Websuche oder Datei-Uploads.

Wichtige Unterscheidungsmerkmale sind Unterstützung für langen Kontext, Echtzeit-Websuche, Sprachein-/-ausgabe, multimodales (Bild-)Verständnis und Programmierfähigkeiten. Einige glänzen beim Coding (CodeStral, o1), andere bei der Genauigkeit im Gespräch (Claude) oder beim visuellen Verständnis (Gemini).

Artificial Analysis veröffentlicht detaillierte LLM-Benchmarks mit Kennzahlen zu Latenz, Kosten und Qualität. Unser Chatbot-Vergleich verlinkt auf Daten auf Modellebene für eine tiefere Analyse. LLM-Benchmarks ansehen