This dataset is combined and deduplicated version of coco-2014 and coco-2017 datasets for object detection. The labels are in Turkish and the dataset is in an instruction-tuning format with separate columns for prompts and completion labels.
For the bounding boxes, a similar annotation scheme to that of PaliGemma annotation is used. That is,
The bounding box coordinates are in the form of special <loc[value]> tokens, where value is a number that represents a normalized coordinate. Each… See the full description on the dataset page:
https://huggingface.co/datasets/ucsahin/COCO-OD-TR-Single-Objects-v2.