EVE-Hallucination is a specialized dataset designed to evaluate language models' tendency to hallucinate (generate factually incorrect or unsupported information) in the Earth Observation (EO) domain. Unlike typical QA datasets that focus on correctness, this dataset contains deliberately hallucinated answers with detailed annotations marking which portions of the text are hallucinated.
This dataset is crucial for developing and evaluating hallucination detection… See the full description on the dataset page:
https://huggingface.co/datasets/eve-esa/hallucination.