Comparación de chatbots de IA

Agentes conversacionales integrados en sitios web, aplicaciones o canales de mensajería que responden preguntas o realizan tareas acotadas en una interfaz de chat, normalmente sobre FAQs, flujos simples o conocimiento específico del dominio, más que sobre tareas de larga duración.

Para comparar modelos de lenguaje, consulta nuestros benchmarks de modelos.

We use AI to collect some results

Highlights

Intelligence Index · Higher is better
Feature score · Higher is better · 6 categories tracked
Monthly price (USD) · Standard & premium plans

Compare Plan Options

ProductTierPriceModelIntelligenceInputsMedia GenToolsMemoryMCPConnectorsFeature ScoreAppsPrivacy
ChatGPT Plus
OpenAIOpenAI
Standard
$20/moGPT-5.4 (xhigh)51
ImagePDFExcelVoice
ImageVideoVoiceV2V
WebCodeDataResearch
MemoryHistory
9/104.7/6
iOSAndroidmacOSWindows
Claude Pro
AnthropicAnthropic
Standard
$20/moClaude Opus 4.7 (max)53
ImagePDFExcelVoice
Voice
WebCodeDataResearch
History
3/103.9/6
iOSAndroidmacOSWindows
Google AI Pro
GoogleGoogle
Standard
$20/moGemini 3.1 Pro Preview46
ImagePDFExcelVideoVoice
ImageVideoVoiceV2V
WebCodeDataResearch
MemoryHistory
5/104.5/6
iOSAndroid
Poe Pro
PoePoe
Standard
$20/mo
Claude Opus 4.6 (max)
43
ImagePDFExcelVoice
ImageVideo
Web
History
0/102.0/6
iOSAndroidmacOSWindows
Perplexity Pro
PerplexityPerplexity
Standard
$20/mo
Claude Opus 4.6 (max)
43
ImagePDFExcelVoice
ImageVoice
WebCodeDataResearch
History
2/103/6
iOSAndroidmacOSWindows
Microsoft Copilot Pro
MicrosoftMicrosoft
Standard
$20/moGPT-5.2 (xhigh)42
ImagePDFExcelVoice
ImageVoice
WebCodeDataResearch
History
4/103.2/6
iOSAndroidWindows
SuperGrok
SpaceXAISpaceXAI
Standard
$30/moGrok 4.1 Fast30
ImagePDFExcelVoice
ImageVoice
WebCodeDataResearch
MemoryHistory
2/103.5/6
iOSAndroid
Mistral Le Chat Pro
MistralMistral
Standard
$15/moMagistral Medium 1.217
ImagePDFExcelVoice
ImageVoice
WebCodeDataResearch
MemoryHistory
2/104.5/6
iOSAndroid

Resumen del panorama

Los principales proveedores —Claude, ChatGPT, Gemini y Meta AI— ofrecen fortalezas distintas en razonamiento, calidad de respuesta y funciones. La mayoría tiene planes gratuitos y planes de pago ($15–25/mes) con ventanas de contexto más grandes, búsqueda web, carga de archivos y generación de imágenes. El mercado se ha consolidado alrededor de estos líderes; opciones open-source como Llama y Mistral ganan adopción cuando la privacidad de los datos importa.

Diferenciadores

  • Algunos destacan en programación (CodeStral, o1), otros en precisión conversacional (Claude) o comprensión visual (Gemini).
  • La diferenciación viene de capacidades especializadas —contexto largo, búsqueda en tiempo real, voz y multimodalidad—, no de la calidad conversacional base, que ha alcanzado paridad.

Intelligence

Artificial Analysis Intelligence Index: Max intelligence supported model

Artificial Analysis Intelligence Index · Higher is better

Artificial Analysis Intelligence Index v4.1 includes: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. See Intelligence Index methodology for further details, including a breakdown of each evaluation and how we run them.

Features

Intelligence Index vs. Feature Score

Feature score · Higher is better · 6 categories tracked
Most attractive quadrant

A summary score counting the features offered by chatbots across six main categories (all listed in the full comparison table above):

  • Media Generation: 1 point (image generation: 0.25, video generation: 0.25, voice conversation: 0.25, native voice-to-voice: 0.25)
  • Tools: 1 point (web search: 0.25, code interpreter: 0.25, data analysis: 0.25, deep research: 0.25)
  • Input Capabilities: 1 point (image input: 0.20, PDF input: 0.20, Excel/CSV input: 0.20, video input: 0.20, voice input: 0.20)
  • Memory: 1 point (memory: 0.5, chat history: 0.5)
  • MCP Support: 1 point (Model Context Protocol integration: 1.0)
  • Connectors: 1 point (available integrations and connectors for external services and tools; each connector weighs 0.1 points; see the 10 connectors evaluated in the comparison table above)

The maximum total value for this metric is 6 points.

Artificial Analysis Intelligence Index v4.1 includes: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. See Intelligence Index methodology for further details, including a breakdown of each evaluation and how we run them.

Price

Intelligence Index vs. Price (Paid Plans)

Artificial Analysis Intelligence Index · Monthly price (USD) · Standard & premium plans
Most attractive quadrant

These charts reveal the pricing structure and value proposition of different paid chatbot plans.

Key insights: The market shows a clear pricing hierarchy with three distinct tiers. Premium plans (SuperGrok Heavy, Google AI Ultra, Claude Max, Perplexity Max, ChatGPT Pro) are positioned at ~$200-300/month, offering top-tier capabilities. Standard plans cluster at much more affordable pricing ($15-30/month) with most options around $20/month, providing accessible AI capabilities for broader audiences.

Monthly Price: The monthly subscription cost in USD for standard and premium chatbot plans. Free plans are excluded from this analysis.

Artificial Analysis Intelligence Index v4.1 includes: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. See Intelligence Index methodology for further details, including a breakdown of each evaluation and how we run them.

Monthly Price: Paid Chatbot Plans

Monthly price (USD) · Standard & premium plans

These charts reveal the pricing structure and value proposition of different paid chatbot plans.

Key insights: The market shows a clear pricing hierarchy with three distinct tiers. Premium plans (SuperGrok Heavy, Google AI Ultra, Claude Max, Perplexity Max, ChatGPT Pro) are positioned at ~$200-300/month, offering top-tier capabilities. Standard plans cluster at much more affordable pricing ($15-30/month) with most options around $20/month, providing accessible AI capabilities for broader audiences.

Monthly Price: The monthly subscription cost in USD for standard and premium chatbot plans. Free plans are excluded from this analysis.

Feature Score vs. Price (Paid Plans)

Feature score · Higher is better · 6 categories tracked · Monthly price (USD) · Standard & premium plans
Most attractive quadrant

This scatter plot reveals the feature value proposition of different paid chatbot plans by comparing their comprehensive feature offerings against monthly pricing.

Key insights: The market shows a clear pricing hierarchy with three distinct tiers. Premium plans (SuperGrok Heavy, Google AI Ultra, Claude Max, Perplexity Max, ChatGPT Pro) are positioned at ~$200-300/month, offering top-tier capabilities. Standard plans cluster at much more affordable pricing ($15-30/month) with most options around $20/month, providing accessible AI capabilities for broader audiences. Notably, some standard plans outperform premium alternatives on features, highlighting diverse value propositions across price points.

Monthly Price: The monthly subscription cost in USD for standard and premium chatbot plans. Free plans are excluded from this analysis.

Preguntas frecuentes

Our comparison uses benchmark data including the Intelligence Index, feature scores, context window size, and pricing. You can filter by model, compare plans side by side, and view charts for intelligence vs price and intelligence vs features to find the best fit for your use case.

The Intelligence Index is a composite score based on Artificial Analysis benchmarking that measures reasoning, knowledge, and response quality. Higher scores indicate stronger performance on standardized evaluations. Paid plans typically score higher than free tiers.

Most major providers (Claude, ChatGPT, Gemini, Meta AI) offer free tiers with varying limits. Our comparison table shows feature scores, context windows, and capabilities for each plan. Free tiers often have smaller context windows and fewer advanced features like web search or file uploads.

Key differentiators include long context support, real-time web search, voice input/output, multimodal (image) understanding, and coding capabilities. Some excel at coding (CodeStral, o1), others at conversational accuracy (Claude) or visual understanding (Gemini).

Artificial Analysis publishes detailed LLM benchmarks including latency, cost, and quality metrics. Our chatbot comparison links to model-level data for deeper analysis. View LLM benchmarks