BLUFF is a comprehensive multilingual benchmark for fake news detection spanning 79 languages with over 202K samples. It uniquely covers both high-resource "big-head" (20) and low-resource "long-tail" (59) languages, addressing critical gaps in multilingual disinformation research.
Paper: BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages
Project Page:… See the full description on the dataset page:
https://huggingface.co/datasets/jsl5710/BLUFF.