A Serbian single-speaker speech dataset prepared for Text-to-Speech (TTS) training using the Common Voice dataset formatting style.
This dataset was created for training Serbian TTS models using frameworks such as Coqui TTS and VITS.
The dataset contains high-quality Serbian speech recordings paired with manually created text transcripts written in Cyrillic script.
Language: Serbian… See the full description on the dataset page:
https://huggingface.co/datasets/daremc86/serbian_common_voice.