llm leaderboard

Best AI for video generation

Human preference over generated clips, from a prompt, from a still, and from an edit.

Rumeqo runs these models inside your team rooms. See what each one costs.

58 of 58 ranked models
Ranked models
rankmodelvendorcompositebenchmarkseloeloelo
1
Dreamina Seedance 2.0 720pbytedance/dreamina-seedance-2.0-720p
ByteDance78.43 of 3 benchmarks147814781377
2
Gemini Omni Flashgoogle/gemini-omni-flash
Google74.63 of 3 benchmarks151214621347
3
Minimax H3minimax/minimax-h3
MiniMax71.22 of 3 benchmarks14531476
4
HappyHorse 1.0alibaba/happyhorse-1.0
Alibaba68.33 of 3 benchmarks142814421308
5
Veo 3.1 Audiogoogle/veo-3.1-audio
Google65.82 of 3 benchmarks13641397
6
FLUX.3 Videoblack-forest-labs/flux-3-video
Black Forest Labs64.81 of 3 benchmarks1496
7
Veo 3.1 Audio 1080pgoogle/veo-3.1-audio-1080p
Google64.52 of 3 benchmarks13631390
8
Grok Imagine Video 1.5 720px-ai/grok-imagine-video-1.5-720p
xAI63.91 of 3 benchmarks1462
9
Grok Imagine Video 720px-ai/grok-imagine-video-720p
xAI63.42 of 3 benchmarks13471417
10
Veo 3.1 Fast Audiogoogle/veo-3.1-fast-audio
Google63.32 of 3 benchmarks13621384
11
Muse Videometa/muse-video
Meta63.21 of 3 benchmarks1457
12
Veo 3.1 Fast Audio 1080pgoogle/veo-3.1-fast-audio-1080p
Google61.52 of 3 benchmarks13581371
13
Wan2.7 I2Valibaba/wan2.7-i2v
Alibaba61.51 of 3 benchmarks1427
14
Sora 2 Proopenai/sora-2-pro
OpenAI60.81 of 3 benchmarks1366
15
Veo 3 Fast Audiogoogle/veo-3-fast-audio
Google58.52 of 3 benchmarks13471324
16
Grok Imagine Video 480px-ai/grok-imagine-video-480p
xAI57.51 of 3 benchmarks1383
17
Veo 3 Audiogoogle/veo-3-audio
Google56.72 of 3 benchmarks13391330
18
Vidu Q3 Proshengshu/vidu-q3-pro
Shengshu55.81 of 3 benchmarks1361
19
Wan2.7 T2Valibaba/wan2.7-t2v
Alibaba55.31 of 3 benchmarks1343
20
Kling v3 Prokwaivgi/kling-v3-pro
Kling AI55.01 of 3 benchmarks1359
21
Sora 2openai/sora-2
OpenAI54.51 of 3 benchmarks1340
22
Seedance v1.5 Probytedance/seedance-v1.5-pro
ByteDance53.12 of 3 benchmarks12561307
23
Wan2.6 T2Valibaba/wan2.6-t2v
Alibaba52.91 of 3 benchmarks1333
24
Wan2.5 I2V Previewalibaba/wan2.5-i2v-preview
Alibaba52.61 of 3 benchmarks1322
25
Wan2.6 I2Valibaba/wan2.6-i2v
Alibaba51.81 of 3 benchmarks1311
26
Wan2.5 T2V Previewalibaba/wan2.5-t2v-preview
Alibaba50.51 of 3 benchmarks1249
27
Pixverse v5.6pixverse/pixverse-v5.6
PixVerse50.12 of 3 benchmarks12401299
28
Grok Imagine Videox-ai/grok-imagine-video
SpaceXAI48.91 of 3 benchmarks1264
29
Veo 3google/veo-3
Google48.22 of 3 benchmarks12531256
30
Gen-4.5runway/gen-4.5
Runway48.11 of 3 benchmarks1223
31
Kling 2.6 Prokwaivgi/kling-2.6-pro
Kling AI47.72 of 3 benchmarks12171293
32
Kling 2.5 Turbo 1080pkwaivgi/kling-2.5-turbo-1080p
Kling AI47.72 of 3 benchmarks12191274
33
Veo 3 Fastgoogle/veo-3-fast
Google47.62 of 3 benchmarks12481256
34
Vidu Q2 Turboshengshu/vidu-q2-turbo
Shengshu44.41 of 3 benchmarks1242
35
Hailuo 2.3minimax/hailuo-2.3
MiniMax44.12 of 3 benchmarks12031260
36
Seedance v1 Probytedance/seedance-v1-pro
ByteDance43.52 of 3 benchmarks11901272
37
Kling O3 Prokwaivgi/kling-o3-pro
Kling AI43.41 of 3 benchmarks1251
38
Kling v2.1 Standardkwaivgi/kling-v2.1-standard
Kling AI42.01 of 3 benchmarks1227
39
Kandinsky 5.0 T2V Prosber/kandinsky-5.0-t2v-pro
Sber AI41.01 of 3 benchmarks1172
40
Hailuo 02 Prominimax/hailuo-02-pro
MiniMax40.42 of 3 benchmarks11981227
41
Ray 3luma/ray-3
Luma AI40.42 of 3 benchmarks12051225
42
Vidu Q2 Proshengshu/vidu-q2-pro
Shengshu39.61 of 3 benchmarks1222
43
Kling O1 Prokwaivgi/kling-o1-pro
Kling AI38.52 of 3 benchmarks12051203
44
Hailuo 02 Fastminimax/hailuo-02-fast
MiniMax37.91 of 3 benchmarks1193
45
Kling v2.1 Masterkwaivgi/kling-v2.1-master
Kling AI37.52 of 3 benchmarks11621234
46
Hailuo 02 Standardminimax/hailuo-02-standard
MiniMax37.42 of 3 benchmarks11801222
47
Kandinsky 5.0 T2V Litesber/kandinsky-5.0-t2v-lite
Sber AI36.21 of 3 benchmarks1113
48
Hunyuan Video 1.5tencent/hunyuan-video-1.5
Tencent35.02 of 3 benchmarks11691196
49
Soraopenai/sora
OpenAI34.61 of 3 benchmarks1069
50
Runway Gen4 Turborunway/runway-gen4-turbo
Runway33.11 of 3 benchmarks1051
51
Mochi v1genmo/mochi-v1
Genmo32.21 of 3 benchmarks1005
52
Runway Gen4 Alephrunway/runway-gen4-aleph
Runway32.21 of 3 benchmarks1194
53
Veo 2google/veo-2
Google32.02 of 3 benchmarks11641164
54
Wan v2.2 A14balibaba/wan-v2.2-a14b
Alibaba30.82 of 3 benchmarks11311169
55
Seedance v1 Litebytedance/seedance-v1-lite
ByteDance30.22 of 3 benchmarks11121184
56
Ltx 2 19Blightricks/ltx-2-19b
Lightricks30.22 of 3 benchmarks11511151
57
Ray2luma/ray2
Luma AI26.62 of 3 benchmarks10641106
58
Pika v2.2pika/pika-v2.2
Pika24.82 of 3 benchmarks1008996
How this ranks

Every benchmark value becomes a percentile among the models that have it, so accuracy scores, Elo ratings and word error rates compare without hand-tuned scaling. Metrics where lower is better are inverted first. Raw values are never summed or averaged across benchmarks. A model's mean percentile is then shrunk toward the mean of the models that were broadly benchmarked, so a model tested twice cannot outrank a broadly tested one on two lucky results. Turning a data source off runs that same ranking code again in your browser over the sources you left on.

A model scored on fewer than 2 of the 3 ranked benchmarks in this category still ranks here, on the benchmarks it does have, and its row carries a partial coverage mark. On an equal score it sits under the model that earned the same number across more of the board.

Data sources

Turn a source off to drop every benchmark it feeds and rank the board again from what is left, in your browser. Turn them all off and the table has nothing to rank. Your choice follows you across the leaderboard pages.

  • LMArenaCC BY 4.0

    Arena ratings by LMArena, from the public leaderboard dataset.