This dataset consists of 28 short audio files in WAV format. Each file contains one sentence in Mandarin, read by a 20-year-old male Chinese speaker. The total duration of the dataset is approximately 3 minutes and 52 seconds. All sentences were recorded and split using Praat.
Issue 1:During recording, the input volume was too high and could not be resolved by lowering the speaker’s voice.
Solution 1:Addjust the microphone… See the full description on the dataset page:
https://huggingface.co/datasets/eduhk-compling/11537448_LiYuzhe.