Text to Video Leaderboard (With Audio)Artificial Analysis

Added to the leaderboard in the last month:

Wan2.7-260612, Gemini Omni Flash, HappyHorse-1.1

Range
Creator
Model
Elo
95% CI
Samples
Released
API Pricing 1
11
Google logoGoogle
Gemini Omni Flash
1,244-11/113,741May 2026$6.00 /min
22
ByteDance Seed logoByteDance Seed
Dreamina Seedance 2.0 720p
1,227-8/810,617Mar 2026$9.07 /min
33
Alibaba logoAlibaba
Wan2.7-260612
1,163-10/103,499Jun 2026$9.00 /min
44
Alibaba-ATH logoAlibaba-ATH
HappyHorse-1.1
1,149-10/103,929Jun 2026$9.90 /min
55
Alibaba-ATH logoAlibaba-ATH
HappyHorse-1.0
1,127-8/87,104Apr 2026$13.20 /min
66-8
KlingAI logoKlingAI
Kling 3.0 1080p (Pro)
1,111-8/89,343Feb 2026$20.16 /min
76-9
Skywork AI logoSkywork AI
SkyReels V4
1,108-9/94,230Mar 2026$21.00 /min
86-9
Alibaba logoAlibaba
Wan 2.7
1,106-10/103,751Apr 2026$9.00 /min
98-12
KlingAI logoKlingAI
Kling 3.0 720p (Standard)
1,100-8/89,271Feb 2026$15.12 /min
109-14
Google logoGoogle
Veo 3.1
1,096-8/87,569Jan 2026$24.00 /min
119-14
KlingAI logoKlingAI
Kling 3.0 Omni 1080p (Pro)
1,096-8/87,409Feb 2026$16.80 /min
129-14
KlingAI logoKlingAI
Kling 3.0 Omni 720p (Standard)
1,094-8/87,307Feb 2026$13.44 /min
1310-14
Google logoGoogle
Veo 3.1 Fast
1,091-8/88,699Jan 2026$9.00 /min
1410-15
Google logoGoogle
Veo 3.1 Lite
1,090-8/86,964Mar 2026$4.80 /min
1514-15
Vidu logoVidu
Vidu Q3 Pro
1,082-8/89,539Jan 2026$9.60 /min
1616-17
PixVerse logoPixVerse
PixVerse V6
1,072-8/88,990Mar 2026$6.90 /min
1716-17
SpaceXAI logoSpaceXAI
grok-imagine-video
1,067-8/89,356Jan 2026$4.20 /min
1818
Alibaba logoAlibaba
Wan 2.6
1,027-8/86,276Dec 2025$9.00 /min
1919
ByteDance Seed logoByteDance Seed
Seedance 1.5 pro
1,0000/07,029Dec 2025$11.86 /min
2020
KlingAI logoKlingAI
Kling 2.6 Pro (January)
987-8/87,291Jan 2026$8.40 /min
2121
Lightricks logoLightricks
LTX-2.3 FastHugging FaceOpen Weights
976-8/87,831Mar 2026$2.40 /min
2222
Lightricks logoLightricks
LTX-2.3 ProHugging FaceOpen Weights
962-8/87,833Mar 2026$4.80 /min
2323-24
Lightricks logoLightricks
LTX-2 FastHugging FaceOpen Weights
950-9/95,922Oct 2025$2.40 /min
2423-24
PixVerse logoPixVerse
PixVerse V5.6
948-9/96,051Feb 2026Coming soon
2525
Lightricks logoLightricks
LTX-2 ProHugging FaceOpen Weights
921-9/95,940Oct 2025$3.60 /min
2626
Sapiens AI logoSapiens AI
Agnes-Video-V2.0
909-10/104,208May 2026$0.30 /min

1 API Pricing reflects the cost to generate 1 minute of 1080p video on the model creator's API at the model's default settings

Frequently Asked Questions

Gemini Omni Flash currently leads among Text to Video models with audio output in the Artificial Analysis Text to Video Arena with an Elo score of 1244.

The top Text to Video models with audio by Elo rating are: 1. Gemini Omni Flash (Elo 1244), 2. Dreamina Seedance 2.0 720p (Elo 1227), 3. Wan2.7-260612 (Elo 1163), 4. HappyHorse-1.1 (Elo 1149), 5. HappyHorse-1.0 (Elo 1127). Rankings are based on blind user votes in the Artificial Analysis Video Arena.

LTX-2.3 Fast currently leads among open weights Text to Video models with audio in the Artificial Analysis Text to Video Arena with an Elo score of 976, followed by LTX-2.3 Pro (Elo 962) and LTX-2 Fast (Elo 950).

Gemini Omni Flash currently leads the Artificial Analysis Text to Video Arena (without audio) with an Elo score of 1326.

The top Text to Video models without audio by Elo rating are: 1. Gemini Omni Flash (Elo 1326), 2. HappyHorse-1.0 (Elo 1288), 3. HappyHorse-1.1 (Elo 1273), 4. Dreamina Seedance 2.0 720p (Elo 1272), 5. Kling 3.0 1080p (Pro) (Elo 1244). Rankings are based on blind user votes in the Artificial Analysis Video Arena.

LTX-2 Pro currently leads among open weights Text to Video models without audio in the Artificial Analysis Text to Video Arena with an Elo score of 1129, followed by LTX-2.3 Fast (Elo 1127) and LTX-2 Fast (Elo 1123).

Models are ranked using an Elo rating system derived from user votes in blind comparisons in the Artificial Analysis Video Arena. Users compare two videos generated from the same text prompt without knowing which model created each video. Higher Elo scores indicate a model is preferred more often. Vote in the Artificial Analysis Video Arena