WildHallucinations is designed for evaluating the factuality of LLMs.
Its core idea is to prompt LLMs to generate and fact-check information about a diverse set of entities.
WildHallucinations consists of 7917 entities extracted from WildChat and a knowledge source.
These entities come from English conversations that are marked as non-toxic.
As described in the main paper, we apply extensive filtering for quality control,
especially for removing entities with more than one meaning.
The… See the full description on the dataset page:
https://huggingface.co/datasets/wentingzhao/WildHallucinations.