A high-quality Lithuanian text-to-speech (TTS) dataset created under the LIEPA project at Vilniaus universitetas (Vilnius University). This dataset serves as a benchmark for Lithuanian speech synthesis, featuring professional recordings from four distinct speakers.
The dataset contains approximately 3 hours of speech per speaker, totaling around 12 hours. It was originally developed for the study Lietuviško balso sintezatorių kokybės… See the full description on the dataset page:
https://huggingface.co/datasets/meldynamics/liepa-tts.