dataset_ar.jsonl is a synthetic Arabic-language fact-checking dataset generated by a two-agent pipeline (Agent A + Agent B). It is designed for training and evaluating language models on the task of validating factual claims in Arabic text.
Format: JSON Lines (one record per line)
Size: 10,008 records / 60,250 items
Language: Arabic MSA (Modern Standard Arabic) — 100%
Total records
10… See the full description on the dataset page:
https://huggingface.co/datasets/fadhel-alobaidi/AraFact.