RefOI-TLHF is a companion dataset to RefOI, developed as part of the study "Vision-Language Models Are Not Pragmatically Competent in Referring Expression Generation."
This dataset focuses on token-level human feedback: for each referring expression—produced by either a human or a model—we annotate the minimal informative span that enables successful identification of the… See the full description on the dataset page:
https://huggingface.co/datasets/Seed42Lab/RefOI-TLHF.