Notice: We added new data and restructured the dataset on 31st October 2024 (GMT+7)
Changes:
Group unique texts together
The annotators of a text are now set as a list of annotator_id. Each respective column is a list of the same size of annotators_id.
Added Polarized column
Notice 2: We rename the dataset from IndoToxic2024 to IndoDiscourse
A Multi-Labeled Dataset for Indonesian Discourse: Examining Toxicity, Polarization, and Demographics Information… See the full description on the dataset page: https://huggingface.co/datasets/JUU198123/IndoDiscourse.