Image to Video Leaderboard (With Audio)Artificial Analysis

Added to the leaderboard in the last month:

Gemini Omni Flash, HappyHorse-1.1

Range
Creator
Model
Elo
95% CI
Samples
Released
API Pricing 1
11-2
Google logoGoogle
Gemini Omni Flash
1,204-11/112,905May 2026$6.00 /min
21-2
ByteDance Seed logoByteDance Seed
Dreamina Seedance 2.0 720p
1,197-8/88,913Mar 2026$9.07 /min
33-4
SpaceXAI logoSpaceXAI
grok-imagine-video-1.5-preview
1,117-10/103,745May 2026$8.40 /min
43-4
Alibaba-ATH logoAlibaba-ATH
HappyHorse-1.1
1,110-10/103,109Jun 2026$9.90 /min
55-7
Alibaba logoAlibaba
Wan 2.7
1,097-10/103,173Apr 2026$9.00 /min
65-7
Alibaba-ATH logoAlibaba-ATH
HappyHorse-1.0
1,092-9/96,501Apr 2026$13.20 /min
76-9
Google logoGoogle
Veo 3.1
1,088-8/86,958Jan 2026$24.00 /min
86-11
Skywork AI logoSkywork AI
SkyReels V4
1,082-10/103,474Mar 2026$21.00 /min
97-10
SpaceXAI logoSpaceXAI
grok-imagine-video
1,082-8/88,339Jan 2026$4.20 /min
108-12
Google logoGoogle
Veo 3.1 Fast
1,079-8/87,401Jan 2026$9.00 /min
1110-13
KlingAI logoKlingAI
Kling 3.0 1080p (Pro)
1,073-8/87,906Feb 2026$20.16 /min
1210-14
PixVerse logoPixVerse
PixVerse V6
1,072-8/88,307Mar 2026$6.90 /min
1311-16
KlingAI logoKlingAI
Kling 3.0 720p (Standard)
1,068-8/87,963Feb 2026$15.60 /min
1412-16
Vidu logoVidu
Vidu Q3 Pro
1,065-8/88,542Jan 2026$9.60 /min
1513-16
KlingAI logoKlingAI
Kling 3.0 Omni 1080p (Pro)
1,062-8/86,243Feb 2026$16.80 /min
1613-16
Google logoGoogle
Veo 3.1 Lite
1,061-9/96,212Mar 2026$4.80 /min
1717
KlingAI logoKlingAI
Kling 3.0 Omni 720p (Standard)
1,052-8/86,252Feb 2026$13.44 /min
1818-19
KlingAI logoKlingAI
Kling 2.6 Pro (January)
1,006-8/86,953Jan 2026$8.40 /min
1919
ByteDance Seed logoByteDance Seed
Seedance 1.5 pro
1,0000/06,323Dec 2025$11.86 /min
2020-22
Lightricks logoLightricks
LTX-2.3 ProHugging FaceOpen Weights
957-9/97,089Mar 2026$4.80 /min
2120-22
Lightricks logoLightricks
LTX-2.3 FastHugging FaceOpen Weights
956-9/96,991Mar 2026$2.40 /min
2220-22
PixVerse logoPixVerse
PixVerse V5.6
954-9/95,716Feb 2026Coming soon
2323
Lightricks logoLightricks
LTX-2 FastHugging FaceOpen Weights
939-9/95,611Oct 2025$2.40 /min
2424
Sapiens AI logoSapiens AI
Agnes-Video-V2.0
925-10/104,686May 2026$0.30 /min
2525
Alibaba logoAlibaba
Wan 2.6
900-9/95,664Dec 2025$9.00 /min
2626
Lightricks logoLightricks
LTX-2 ProHugging FaceOpen Weights
880-10/105,603Oct 2025$3.60 /min

1 API Pricing reflects the cost to generate 1 minute of 1080p video on the model creator's API at the model's default settings

Frequently Asked Questions

Gemini Omni Flash currently leads among Image to Video models with audio output in the Artificial Analysis Image to Video Arena with an Elo score of 1204.

The top Image to Video models with audio by Elo rating are: 1. Gemini Omni Flash (Elo 1204), 2. Dreamina Seedance 2.0 720p (Elo 1197), 3. grok-imagine-video-1.5-preview (Elo 1117), 4. HappyHorse-1.1 (Elo 1110), 5. Wan 2.7 (Elo 1097). Rankings are based on blind user votes in the Artificial Analysis Video Arena.

LTX-2.3 Pro currently leads among open weights Image to Video models with audio in the Artificial Analysis Image to Video Arena with an Elo score of 957, followed by LTX-2.3 Fast (Elo 956) and LTX-2 Fast (Elo 939).

Gemini Omni Flash currently leads the Artificial Analysis Image to Video Arena (without audio) with an Elo score of 1374.

The top Image to Video models without audio by Elo rating are: 1. Gemini Omni Flash (Elo 1374), 2. Dreamina Seedance 2.0 720p (Elo 1346), 3. grok-imagine-video-1.5-preview (Elo 1330), 4. PixVerse V6 (Elo 1329), 5. grok-imagine-video (Elo 1328). Rankings are based on blind user votes in the Artificial Analysis Video Arena.

Cosmos3-Super-Image2Video currently leads among open weights Image to Video models without audio in the Artificial Analysis Image to Video Arena with an Elo score of 1252, followed by LTX-2 Pro (Elo 1192) and LTX-2 Fast (Elo 1183).

Text to Video models generate videos from text descriptions alone, while Image to Video models take an existing image as input and animate or extend it into a video. This allows for more control over the visual content and style of the generated video.

Models are ranked using an Elo rating system derived from user votes in blind comparisons. Users compare videos generated from the same input image and choose the result they prefer. Higher Elo scores indicate a model is preferred more often. Vote in the Artificial Analysis Video Arena