SciFact-Fa is a Persian (Farsi) dataset designed for the Retrieval task, with a focus on scientific fact verification. It is a translated version of the original English SciFact dataset used in the BEIR benchmark and is part of the FaMTEB (Farsi Massive Text Embedding Benchmark) under the BEIR-Fa collection.
Language(s): Persian (Farsi)
Task(s): Retrieval (Scientific Fact Verification, Evidence Retrieval)
Source: Translated from the English SciFact dataset using… See the full description on the dataset page:
https://huggingface.co/datasets/MCINext/scifact-fa.