FeruzaSpeech is a read speech dataset
of the Uzbek language, transcribed in both Cyrillic
and Latin alphabets, freely available for academic research purposes.
It includes 60 hours of high-quality recordings
from a single native female speaker from Tashkent, Uzbekistan.
ICNLSPConference:
https://www.youtube.com/watch?v=9whj9yzI_s4&ab_channel=ICNLSPConference
Paper:
https://arxiv.org/abs/2410.00035
audio text_latin… See the full description on the dataset page:
https://huggingface.co/datasets/k2speech/FeruzaSpeech.