InclusionAI: inteligencia, rendimiento y precio de sus modelos

InclusionAI
InclusionAI

Este análisis está pensado para ayudarte a elegir el mejor modelo ofrecido por InclusionAI para tu caso de uso.

Más inteligente

Updated
#1
Ling 3.0 FlashLing 3.0 Flash
25
#2
Ling-3.0-flash-VLLing-3.0-flash-VL
25
#3
Ling-3.0-flash-FinLing-3.0-flash-Fin
23
#4
Ring-2.6-1TRing-2.6-1T
17
#5
Ling 3.0 TinyLing 3.0 Tiny
15

Índice de inteligencia

5 modelos en total

Más rápido

#1
Ling 3.0 FlashLing 3.0 Flash
335 t/s
#2
Ling-3.0-flash-FinLing-3.0-flash-Fin
166 t/s
#3
Ling-3.0-flash-VLLing-3.0-flash-VL
145 t/s
#4
Ring-2.6-1TRing-2.6-1T
117 t/s
#5
Ling 3.0 TinyLing 3.0 Tiny
49 t/s

Velocidad de salida

5 modelos en total

Menor precio

#1
Ling 3.0 FlashLing 3.0 Flash
$0.05
#2
Ling-3.0-flash-VLLing-3.0-flash-VL
$0.05
#3
Ling-3.0-flash-FinLing-3.0-flash-Fin
$0.05
#4
Ring-2.6-1TRing-2.6-1T
$0.52

Precio combinado (por 1M de tokens)

5 modelos en total

InclusionAI ofrece 5 modelos, cada uno con distintas características de inteligencia, rendimiento y precio. A continuación se comparan las métricas clave entre modelos.

  • En inteligencia, los mejores modelos en InclusionAI son Ling 3.0 Flash (25), Ling-3.0-flash-VL (25) y Ling-3.0-flash-Fin (23).
  • En velocidad de salida, los modelos más rápidos son Ling 3.0 Flash (335 t/s), Ling-3.0-flash-Fin (166 t/s) y Ling-3.0-flash-VL (145 t/s). La velocidad varía significativamente entre modelos, con una diferencia del 588% entre el más rápido y el más lento.
  • En latencia, Ling 3.0 Flash (9.11s), Ling-3.0-flash-Fin (13.79s) y Ling-3.0-flash-VL (15.67s) ofrecen el menor tiempo hasta el primer token de respuesta.
  • En precios, Ling 3.0 Flash ($0.05), Ling-3.0-flash-VL ($0.05) y Ling-3.0-flash-Fin ($0.05) ofrecen los menores precios combinados por 1M de tokens. Los precios varían hasta 10.9x entre modelos.
  • En tamaño de ventana de contexto, Ling 3.0 Flash (262k), Ling-3.0-flash-VL (262k) y Ling-3.0-flash-Fin (262k) admiten las ventanas de contexto más grandes en InclusionAI.
  • Ling 3.0 Flash destaca como el líder general en InclusionAI, al ubicarse primero en inteligencia, velocidad y precios.

Aspectos destacados

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Evaluaciones de inteligencia

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

No data available

Agentic business operations

No data available

Agentic scientific research workflows in a terminal

No data available

Quantitative analysis on spreadsheets & documents

Agentic tool use

Kubernetes incident root-cause analysis

No data available

Visual reasoning

Medical long context reasoning

No data available

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Ventana de contexto

Context Window

Context window: tokens limit · Higher is better

Precios

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Resumen de rendimiento

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant

Velocidad

Medida por la velocidad de salida (tokens por segundo)

Output Speed

Output tokens per second · Higher is better

Latencia

Medida por el tiempo (segundos) hasta el primer token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

Tiempo de respuesta de extremo a extremo

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant

Análisis adicional
Logo de InclusionAI
Ling 3.0 Flash
262k
Abierto
25*
--
335
3.14
10.60
5.96
Logo de InclusionAI
Ling-3.0-flash-VL
262k
Abierto
25
--
145
1.86
19.13
13.81
Logo de InclusionAI
Ling-3.0-flash-Fin
262k
Abierto
23
--
166
1.73
16.81
12.06
Logo de InclusionAI
Ling-2.6-1T
262k
Abierto
17*
--
--
--
--
--
Logo de InclusionAI
Ring-2.6-1T
262k
Abierto
17
$0.29
117
4.15
25.49
17.08
Logo de InclusionAI
Ling 3.0 Tiny
262k
Abierto
15*
--
49
2.65
53.93
41.03

Definiciones clave

Preguntas frecuentes

Preguntas comunes sobre InclusionAI