Hypa-Voices is an open-source multilingual, multi-modal audio-text dataset from Hypa Intelligence and AfroVoices, with a long-term vision of advancing speech and language technology for under-represented languages. Every record pairs written text with corresponding speech, either in the same language (transcription) or across two languages while carrying the same meaning (translation). In translation records, speech is in src_lang and text is in tgt_lang.
This… See the full description on the dataset page:
https://huggingface.co/datasets/hypaai/Hypa-Voices.