MMVU benchmark maintainers
Multimodal Multi-disciplinary Video Understanding
A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.
ModelsModels with a result.7
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.2026
Result and price
Results
7 results#ModelResultThe score from the source.