This dataset is the Natural Questions (NQ) slice of the UDA (Unstructured Document Analysis) benchmark, packaged for retrieval-oriented evaluation: 2,477 question–answer instances grounded in Wikipedia articles, with gold long and short answers plus document identifiers.
Natural Questions (Kwiatkowski et al., TACL 2019) is a large-scale benchmark built from real anonymized Google Search queries. Annotators were… See the full description on the dataset page: https://huggingface.co/datasets/orgrctera/uda_nq_qa.