Paper: TatarTTS: An Open-Source Text-to-Speech Synthesis Dataset for the Tatar Language
GitHub:
https://github.com/IS2AI/TatarTTS
Description: TatarTTS is an open-source text-to-speech dataset for the Tatar language.
The dataset comprises ~70 hours of transcribed audio recordings, featuring two professional speakers (one male and one female).
Citation:
The project was developed in academic collaboration between ISSAI and Institute of Applied Semiotics of Tatarstan… See the full description on the dataset page:
https://huggingface.co/datasets/issai/TatarTTS.