This dataset comprises 35 WAV audio files. Each file contains one sentence in Mandarin read
by a 19-year-old female Chinese speaker. The dataset's total duration is approximately 3 minutes.
All sentences were recorded and split using Praat.
Issue1: Recording noise due to venue limitations-- Persistent background noise in the dormitory
environment (door opening/closing sounds) adversely affects TTS training quality.… See the full description on the dataset page:
https://huggingface.co/datasets/eduhk-compling/11537541_XiaoJiayi.