This dataset is a modified version of the KILT Benchmark from the paper "KILT: a benchmark for knowledge intensive language tasks". It includes additional top-k retrieval results used in the paper "Chain-of-Retrieval Augmented Generation".
The primary difference is the addition of the context_doc_ids field. This field provides the IDs of the top-k documents retrieved during the… See the full description on the dataset page:
https://huggingface.co/datasets/corag/kilt.