YOLO-format object-detection dataset for Tibetan Document Layout Analysis (TDLA). The dataset contains bounding-box annotations for four layout classes found in Tibetan document page images. It is split into training, validation, and test sets. The train/val split uses iterative multi-label stratification, while the test set is a hand-picked benchmarking set of the most unique page layouts.
Total annotations… See the full description on the dataset page:
https://huggingface.co/datasets/BDRC/TDLA-Training-Dataset.