This dataset provides a collection of Spanish text-to-speech (TTS) audio samples with human naturalness ratings, aimed at advancing research on automatic TTS quality assessment in Spanish.
The dataset contains samples from 52 different speakers, including 12 TTS systems and 6 real human voices. The samples cover multiple Spanish dialects, speaker genders, and speech… See the full description on the dataset page:
https://huggingface.co/datasets/asosawelford/es-TTS-subjective-naturalness.