A single-speaker read speech dataset in Kenyan Swahili, containing approximately 6 hours of prompted recordings from an anonymous male speaker. The dataset was produced as part of CLEAR Global's Gamayun Language Data Kits initiative, which develops open-source language resources for under-resourced languages used in humanitarian contexts.
The sentence set is shared with CLEAR Global's Gamayun Swahili–English parallel text kit. English source… See the full description on the dataset page:
https://huggingface.co/datasets/Nzyoka19/Kenyan-Swahili-Speech.