LithosAI: Models Intelligence, Performance & Price

LithosAI
LithosAI

This analysis is intended to support you in choosing the best model provided by LithosAI for your use-case.

Most Intelligent

Updated
#1
Kimi K3 (max) (ULTRA CHAT)Kimi K3 (max) (ULTRA CHAT)
44
#2
Kimi K3 (max)Kimi K3 (max)
44
#3
Kimi K3 (max) (FAST)Kimi K3 (max) (FAST)
44
#4
Kimi K3 (max) (ULTRA)Kimi K3 (max) (ULTRA)
44
#5
DeepSeek V4.1 Flash (max) (FAST)DeepSeek V4.1 Flash (max) (FAST)
39

Intelligence index

Total 11 models

Fastest

#1
DeepSeek V4.1 Flash (max) (ULTRA CHAT)DeepSeek V4.1 Flash (max) (ULTRA CHAT)
633 t/s
#2
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)
596 t/s
#3
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)
579 t/s
#4
DeepSeek V4.1 Flash (max) (ULTRA)DeepSeek V4.1 Flash (max) (ULTRA)
550 t/s
#5
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
534 t/s

Output speed

Total 11 models

Lowest Price

#1
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
$0.09
#2
DeepSeek V4.1 Flash (max) (FAST)DeepSeek V4.1 Flash (max) (FAST)
$0.15
#3
DeepSeek V4.1 Flash (Non-reasoning) (FAST)DeepSeek V4.1 Flash (Non-reasoning) (FAST)
$0.15
#4
DeepSeek V4.1 Flash (max) (ULTRA)DeepSeek V4.1 Flash (max) (ULTRA)
$0.21
#5
DeepSeek V4.1 Flash (max) (ULTRA CHAT)DeepSeek V4.1 Flash (max) (ULTRA CHAT)
$0.21

Blended price (per 1M tokens)

Total 11 models

LithosAI offers 11 models, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across models.

  • For intelligence, the top models on LithosAI are Kimi K3 (max) (ULTRA CHAT) (44), Kimi K3 (max) (44), and Kimi K3 (max) (FAST) (44).
  • For output speed, the fastest models are DeepSeek V4.1 Flash (max) (ULTRA CHAT) (633 t/s), DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT) (596 t/s), and DeepSeek V4.1 Flash (Non-reasoning) (ULTRA) (579 t/s).
  • For latency, DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT) (0.79s), DeepSeek V4.1 Flash (Non-reasoning) (ULTRA) (1.01s), and DeepSeek V4.1 Flash (Non-reasoning) (FAST) (1.25s) offer the lowest time to first answer token.
  • For pricing, DeepSeek V4.1 Flash (max) ($0.09), DeepSeek V4.1 Flash (max) (FAST) ($0.15), and DeepSeek V4.1 Flash (Non-reasoning) (FAST) ($0.15) offer the lowest blended prices per 1M tokens. Prices vary up to 2.3x across models.
  • For context window size, Kimi K3 (max) (ULTRA CHAT) (1M), Kimi K3 (max) (1M), and Kimi K3 (max) (FAST) (1M) support the largest context windows on LithosAI.

Highlights

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Intelligence Evaluations

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Agentic tool use

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Pricing

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Performance Summary

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Speed

Measured by Output Speed (tokens per second)

Output Speed

Output tokens per second · Higher is better

Latency

Measured by Time (seconds) to First Token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

End-to-End Response Time

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Further Analysis
Kimi logo
Kimi K3 (max) (ULTRA CHAT)
1.05M
Open
44
$3.93
289
0.74
9.40
6.93
Kimi logo
Kimi K3 (max)
1.05M
Open
44
$1.69
219
0.71
12.12
9.13
Kimi logo
Kimi K3 (max) (FAST)
1.05M
Open
44
$2.81
217
0.74
12.25
9.20
Kimi logo
Kimi K3 (max) (ULTRA)
1.05M
Open
44
$3.93
221
0.76
12.07
9.05
DeepSeek logo
DeepSeek V4.1 Flash (max) (FAST)
1.05M
Open
39
$0.27
503
1.43
6.40
3.98
DeepSeek logo
DeepSeek V4.1 Flash (max) (ULTRA)
1.05M
Open
39
$0.38
550
1.32
5.86
3.63
DeepSeek logo
DeepSeek V4.1 Flash (max) (ULTRA CHAT)
1.05M
Open
39
$0.38
633
0.77
4.72
3.16
DeepSeek logo
DeepSeek V4.1 Flash (max)
1.05M
Open
39
$0.16
534
1.09
5.78
3.74
DeepSeek logo
DeepSeek V4.1 Flash (Non-reasoning) (FAST)
1.05M
Open
25
$0.16
485
1.25
2.28
--
DeepSeek logo
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)
1.05M
Open
25
$0.22
579
1.01
1.87
--
DeepSeek logo
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)
1.05M
Open
25
$0.22
596
0.79
1.62
--

Key definitions

Frequently Asked Questions

Common questions about LithosAI