A large-scale multimodal dataset for studying fairness and bias in hate speech detection systems with counterfactual augmentation.
Dataset Description
This dataset contains 18,000 text-image pairs categorized into 8 hate speech classes with varying levels of protected group representation. The dataset was created to evaluate whether Counterfactual Data Augmentation (CDA) introduces or amplifies bias in hate speech detection models.
Key… See the full description on the dataset page: https://huggingface.co/datasets/vs16/counter-hate-dataset.