Multimodal Multi-disciplinary Video Understanding

A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.

ModelsModels with a result.7
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.2026

Result and price

020406080100$0.3$1$3RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

7 results
#ModelResultThe score from the source.
Qwen3.8 Max82.4%
Kimi K2.580.4%
Qwen3.5-27B73.3%