This model is a fine-tuned version of
xlm-roberta-base on the None dataset.
It achieves the following results on the evaluation set:
Version: 1.0
Author: Amaan Shaikh (amaan00z)
Language: Hinglish
Labels:
0 = not_sarcastic
1 = sarcastic
✅ MUStARD Hinglish Dialogues
✅ Swami et al. Hinglish Twitter dataset
✅ 4,000 generated Gen-Z + political + meme sarcasm samples
✅ 2,500 real non-sarcastic Hinglish social-media samples
✅ Reddit & general Hinglish sarcasm templates
✅ Cleaned + deduplicated final dataset: 9,594 samples