SarcNet is a novel benchmark for multilingual and multimodal sarcasm detection, introduced at LREC-COLING 2024. It addresses the limitations of single-language datasets by providing 3,335 image-text pair samples in both English and Chinese.
In contrast to traditional datasets that use a single unified label, SarcNet employs a separated annotation schema. Each sample is distinctly labeled across… See the full description on the dataset page:
https://huggingface.co/datasets/alita9/sarcnet.