This dataset is a collection of social media posts that have been augmented using a large language model (GPT-4o). The original dataset was sourced from the paper Multilingual Previously Fact-Checked Claim Retrieval by Matúš Pikuliak et al. (2023). You can access the original dataset from here. The dataset is used for improving the ability to comprehend content across multiple languages by integrating… See the full description on the dataset page: https://huggingface.co/datasets/MultiMind-SemEval2025/Augmented_MultiClaim_FactCheck_Retrieval.