Daniel Cores*,
Michael Dorkenwald*,
Manuel Mucientes,
Cees G. M. Snoek,
Yuki M. Asano
*Equal contribution.
23 December 2024: Please redownload the dataset, as the Unexpected Action labels have been updated.
TVBench is a new benchmark specifically created to evaluate temporal understanding in video QA. We identified three main issues in existing datasets: (i) static information from single… See the full description on the dataset page:
https://huggingface.co/datasets/FunAILab/TVBench.