A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.
Reference2026Active4 models
Data verified 27 Jul 2026 · Methodology 1.8.0
Benchmark score on Multimodal Multi-disciplinary Video Understanding
About Multimodal Multi-disciplinary Video Understanding
Definition and scoring
Organisation
MMVU benchmark maintainers
Category
Multimodal
Version
2026
Direction
higher is better
Ranking use
Reference
Contamination risk
Unknown
A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content. Results stay tied to the exact model variant and evaluation system. Multiple systems for the same model use the best published score on this page; overall Lumina scoring uses the median of ranking-eligible rows.
What does Multimodal Multi-disciplinary Video Understanding measure?
A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.
Which model scores highest on Multimodal Multi-disciplinary Video Understanding?
Kimi K2.5 by Moonshot AI currently leads with 80.4%.
How many models are evaluated on Multimodal Multi-disciplinary Video Understanding?
4 models in the LuminaBench cohort have a qualifying score on this benchmark.
Does this affect overall Lumina rank?
No. This benchmark is display-only and does not enter the overall Lumina composite.