Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
acholi-speech-data – Dataset by Professor | AlphaNeural AI
You can deploy this model and start earning money today!
Professor
/
acholi-speech-data
like
0
text-to-speech
automatic-speech-recognition
monolingual
ach
cc-by-4.0
1K<n<10K
us
acholi
luo
speech
tts
asr
african-languages
low-resource
Views
No views yet
Model card
Files and Versions
Community
API
Acholi Speech Data (Pooled)
A ~29.7-hour Acholi speech corpus, pooling a crowdsourced ASR config with a studio-quality TTS config. Part of the AfroNet multi-language TTS data effort.
Source
WAXAL (google/WaxalNLP), two configs pooled together:
ach_asr — crowdsourced, image-prompted speech, many speakers. 4,401 clips, 24.8h. ach_tts — studio-quality read speech. 1,442 clips, 4.9h.
train+validation+test splits are pooled together across both configs (intentional… See the full description on the dataset page:
https://huggingface.co/datasets/Professor/acholi-speech-data
.