The dataset contains 4,050 hours of first-person videos for egocentric vision and egocentric tracking. Featuring multimodal data from egocentric views, it includes data annotations and motion capture for extracting 3d poses. It provides detailed 3d objects and 3d scenes using visual data from VR headsets to analyze hands motions and pose estimations.… See the full description on the dataset page:
https://huggingface.co/datasets/UniDataPro/egocentric-video.