This repo contains data from AI21 Labs' paper Generating Benchmarks for Factuality Evaluation of Language Models.
NEWS-FACTOR: Based on Reuters articles extracted from The RefinedWeb Dataset. The dataset consists of 1036 examples.
The benchmark is derived from The RefinedWeb Dataset. The public extract is made available under an ODC-By 1.0 license; users should also abide to the CommonCrawl ToU:
https://commoncrawl.org/terms-of-use/.
Cite:
@article{muhlgay2023generating,
title={Generating… See the full description on the dataset page:
https://huggingface.co/datasets/mansaripo/NEWS-FACTOR.