Perception-Bench is a benchmark for evaluating the long-form response of a VLM (Vision Language Model) across various domains of images, and it is a held-out test
set of the Perception-Collection
Languages
English
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/prometheus-eval/Perception-Bench.