Existing works on underrepresented languages mostly focus
on machine translation and simple language understanding tasks,
such as sentiment analysis and topic classification.
More complex tasks such as open-domain dialogue and dialogue summarization,
are still left behind for these underrepresented languages which leads
to a poor evaluation suite for assessing the capability of
large language models (LMs) in these underrepresented… See the full description on the dataset page: https://huggingface.co/datasets/prosa-text/nusa-dialogue.