gpt.college

Video Ranking

Signals for video understanding and temporal multimodal reasoning.

Main rankings require at least 1 of 1 included sourced benchmark rows. Within the main ranking, rows are sorted by average normalized score; coverage is shown separately and used only as a tie-breaker. Missing benchmark data is not treated as zero; models below the threshold appear under limited evidence.
Rank Model Provider Score Coverage Source Source Type Verified Last Updated

% accuracy

Video-MME

A video understanding benchmark for multimodal models.