This dataset provides 32.7 hours of English speech recordings (7,592 samples) from 44 Rwandan speakers living with speech impairments. The participants represent a limited diversity of speech patterns, mostly stuttering, and a few examples of Dysarthria, Dysphonia, and Phonological disorders.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker or phrase level. All speech recordings of this datasets have been… See the full description on the dataset page:
https://huggingface.co/datasets/cdli/rwandan_english_nonstandard_speech_v1.0.