Dataset Card for the Synthetic Swahili-English Helpline Translation Dataset
Dataset Details
Dataset Description
This dataset contains synthetic parallel Swahili-English translations from Tanzanian child helpline conversations, designed for training and evaluating neural machine translation (NMT) models. The dataset addresses the critical need for high-quality translation systems in child protection services across East Africa, where multilingual support is… See the full description on the dataset page: https://huggingface.co/datasets/openchs/synthetic-helpline-sw-en-translation-v1.