[Paper]
Our code and dataset will be released soon.
@article{han2024videoespresso,
title={VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection},
author={Han, Songhao and Huang, Wei and Shi, Hairong and Zhuo, Le and Su, Xiu and Zhang, Shifeng and Zhou, Xu and Qi, Xiaojuan and Liao, Yue and Liu, Si},
journal={arXiv… See the full description on the dataset page:
https://huggingface.co/datasets/hshjerry0315/VideoEspresso-Test.