This dataset contains Levantine Arabic tweets labeled for Hate Speech and Abusive language. It is used to train the model: amitca71/marabert2-levantine-toxic-model
Dataset Structure
text: The tweet content (Arabic).
label: The classification label.
0: Abusive
1: Normal
2: Hate
Citation
Mulki, H., et al. (2019). "L-HSAB: A Levantine Twitter Dataset for Hate Speech and Abusive Language."