This dataset contains extracted structured annotations for the OpenDV-YouTube
driving-video dataset released by OpenDriveLab as part of
DriveAGI, together with the
OpenDV-YouTube-Language
metadata. It was collected and generated as part of the
MAD project.
The repository includes:
car skeleton keypoints extracted with OpenPifPaf;
lane skeleton keypoints extracted with OpenPifPaf;
pedestrian whole-body keypoints extracted with DWPose;
image captions… See the full description on the dataset page:
https://huggingface.co/datasets/AhmadRH/OpenDV_Poses.