This is the Zulu language split of the NCHLT speech corpus (nchlt-clean split). It containes 56 hours and 14 minutes hours of isiZulu speech from 210 (98 Male; 112 Female) native speakers of the language. Each recording consists of a speaker reading three words. These words were sourced from South African government websites. The full corpus can be downloaded from the SADILAR website.
For a full description of the corpus, see Barnard et al. 2014.
Citation… See the full description on the dataset page: https://huggingface.co/datasets/aconeil/nchlt.