digital_graphics: 0
digital_presentation: 1
digital_screencast: 2
digital_videogame: 3
footage_multi_speaker: 4
footage_scene: 5
footage_talkinghead: 6
The dataset follows the YOLO format:
dataset/
├── images/
│ ├── train/
│ └── val/
├── labels/
│ ├── train/
│ └── val/
└──… See the full description on the dataset page:
https://huggingface.co/datasets/zigg-ai/content-regions-distrib-yolo.