This is speech dataset of Bemba language. This dataset was acquired from (BembaSpeech)[
https://github.com/csikasote/BembaSpeech/tree/master].
BembaSpeech is the speech recognition corpus in Bemba [1].
DatasetDict({
train: Dataset({
features: ['audio', 'sentence'],
num_rows: 12421
})
dev: Dataset({
features: ['audio', 'sentence'],
num_rows: 1700
})
test: Dataset({
features: ['audio'… See the full description on the dataset page:
https://huggingface.co/datasets/kreasof-ai/bemba-speech-csikasote.