BenchLeader

Video-MME

Video understanding across durations, without subtitles.

As of 19 Sept 2026, video-SALMONN 2+ leads Video-MME on BenchLeader with 79.7%, ahead of Gemini 1.5 Pro 001 at 75.0%, across 42 model configurations with a published result.

Published by
Video-MMEdata via Epoch AI Benchmarking Hub
Category
Multimodal
Index weight
Reference only
Models
42
Data as of
19 Sept 2026

CC BY 4.0 — Epoch AI, ‘AI Benchmarking Hub’, epoch.ai/benchmarks. Mirrored boards credit their original publishers.

What the test looks like

Questions about videos from seconds to an hour long, answered without subtitles.

How it is scored

Accuracy, as published by Video-MME.

What to keep in mind

Video-capable models only.

42 of 42
#
1video-SALMONN 2+ByteDance79.7%2025-06-18
2Gemini 1.5 Pro 001Google75.0%45.12024-05-14
3Gemini 1.5 Pro 001 Feb24Google75.0%2024-02-15
4Qwen2.5-VL 72B InstructAlibabaopen ↗73.5%2024-09-19
5InternVL2_5-78BShanghai AI Lab72.1%2024-12-06
6GPT-4oOpenAI71.9%44.32024-11-20
7Qwen2 Vl 72BAlibaba71.2%2024-08-29
8LLaVA-Video-72B-Qwen2OpenBMB70.6%2024-09-02
9Gemini 1.5 Flash 001Google70.3%36.72024-05-23
10ViLAMP-llava-qwenUnknown67.5%2025-05-01
11Oryx-1.5-32BUnknown67.3%2024-10-22
12LLaVA-OneVision 72BOpenBMB66.3%2024-08-06
13VideoLLAMA3-7BUnknown66.2%2024-01-22
14LLaVA-Video-7B-Qwen2OpenBMB65.9%2024-09-02
15LLaVA-Video-7B-Qwen2-TPOOpenBMB65.6%2025-01-19
16VideoChat-Flash-Qwen2-7B_res448Unknown65.3%2025-01-11
17GPT-4o miniOpenAI64.8%36.22024-07-18
18ByteVideoLLM-14BUnknown64.6%2024-10-13
19NVILA 8BNVIDIA64.2%2024-12-10
20LiveCC-7B-InstructUnknown64.1%2024-04-12
21Qwen2 Vl 7BAlibaba63.9%2024-08-29
22MiniCPM-o-2_6OpenBMB63.9%2025-01-12
23VideoLLAMA2-7BUnknown62.4%2024-06-12
24InternVL2-40BShanghai AI Lab61.2%2024-07-08
25Minicpm V 2.6OpenBMB60.9%2024-08-03
26Claude 3.5 SonnetAnthropic60.0%46.62024-06-20
27GPT-4VOpenAI59.9%2023-11-06
28mPLUG-Owl3-7B-241101Unknown59.3%2024-11-26
29VITA-1.5Unknown56.1%2024-12-20
30Video-XL-7BUnknown55.5%2024-10-17
31Video-CCAM-7B-v1.2Unknown53.2%2024-09-29
32long-llava-qwen2-7bUnknown52.9%2024-08-30
33LongVA-7BUnknown52.6%2024-06-13
34Qwen VlmaxAlibaba51.3%2024-01-18
35SliME-Llama3-8BUnknown45.3%2024-06-02
36Chat-Uni-Vi-7B-v1.5 + 100k SG-WVUnknown43.2%2024-06-20
37Qwen VlAlibaba41.1%2023-08-20
38Chat-UniVi-7B-v1.5Unknown40.6%2024-04-23
39sharegpt4video-8bUnknown39.9%2024-05-27
40Video-LLaVA-7BUnknown39.9%2023-11-17
41video_chat2_mistralMistral AI39.5%2023-11-29
42ST-LLMUnknown37.9%2024-03-28

Cite as: BenchLeader, “Video-MME leaderboard”, https://www.benchleader.com/benchmarks/video_mme, data as of 19 Sept 2026.

Video-MME: questions

What does Video-MME measure?
Questions about videos from seconds to an hour long, answered without subtitles. Scores are reported in percent of tasks solved; higher is better.
Which AI model leads Video-MME?
video-SALMONN 2+ leads Video-MME with 79.7% as of 19 Sept 2026, ahead of Gemini 1.5 Pro 001 at 75.0%.
How many models have Video-MME results?
42 model configurations have a Video-MME result on BenchLeader, all taken from Video-MME via Epoch AI Benchmarking Hub.
Who runs Video-MME and how often is it updated?
Video-MME is published by Video-MME. BenchLeader re-reads the published results every morning and records the date each result was published.
Does Video-MME count toward the BenchLeader Index?
No. Video-MME is shown for reference but left out of the composite index.