A ~39.0-hour Lusoga speech corpus, drawn from a single source (WAXAL) and
filtered to only genuinely transcribed audio. Part of the
AfroNet multi-language TTS data
effort.
WAXAL (google/WaxalNLP),
sog_asr config — crowdsourced, image-prompted speech collected via Makerere
University's "Yogera" app (the same pipeline used for WAXAL's Masaaba data). 6,723
clips, 39.0h, source = waxal.
train+validation+test splits are pooled… See the full description on the dataset page:
https://huggingface.co/datasets/Professor/lusoga-speech-data.