SDS KoPub-VDR is a benchmark dataset for Visual Document Retrieval (VDR) in the context of
Korean public documents. It contains real-world government document images paired with natural-language
queries, corresponding answer pages, and ground-truth answers. The dataset is designed to evaluate AI models that
go beyond simple text matching, requiring comprehensive understanding of visual layouts, tables, graphs, and images
to accurately locate relevant… See the full description on the dataset page:
https://huggingface.co/datasets/SamsungSDS-Research/SDS-KoPub-VDR-Benchmark.