PopQA is a large-scale open-domain question answering (QA) dataset, consisting of 14k entity-centric QA pairs. Each question is created by converting a knowledge tuple retrieved from Wikidata using a template. Each question come with the original subject_entitiey, object_entityand relationship_type annotation, as well as Wikipedia monthly page views.
Languages
The dataset contains samples in English only.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/akariasai/PopQA.