VideoMME
VideoMME is a multimodal benchmark tracked on BenchGecko. Video understanding benchmark testing comprehension of video content, temporal reasoning, and scene analysis.
Built from BenchGecko data as of April 9, 2026 · updates when the data changes
VideoMME is a multimodal benchmark tracked on BenchGecko. Video understanding benchmark testing comprehension of video content, temporal reasoning, and scene analysis.
Basic
Video understanding benchmark testing comprehension of video content, temporal reasoning, and scene analysis.
Deep
Scores are reported (%, maximum 100); higher is better. BenchGecko collects VideoMME scores from Epoch AI and refreshes them daily; the leaderboard on this page shows the current top models.
Expert
Compare models only on the same benchmark version and settings. A benchmark separates models well while scores are spread out; when top models bunch together near the maximum, it stops telling them apart.
Depending on why you're here
- ·multimodal benchmark (%, maximum 100)
- ·Source: Epoch AI
- ·Useful if your workload is multimodal
- ·Check the live leaderboard on /benchmark/videomme before choosing a model
- ·Labs cite benchmark results in launch announcements; check the source and version
- ·VideoMME is a test of how good an AI is at multimodal
- ·A higher score means better results on that kind of task