Dataset Description:
This dataset contains a sentiment analysis dataset from Winata et al. (2022).
Data Structure:
The data was used for the project on improving word embeddings with graph knowledge for Low Resource Languages.
Citation:
@misc{winata2022nusax,
title={NusaX: Multilingual Parallel Sentiment Dataset for 10 Indonesian Local Languages},
author={Winata, Genta Indra and Aji, Alham Fikri and Cahyawijaya, Samuel… See the full description on the dataset page:
https://huggingface.co/datasets/DGurgurov/sundanese_sa.