This CSTR VCTK Corpus includes around 44-hours of speech data uttered by 110 English speakers with various accents. Each speaker reads out about 400 sentences, which were selected from a newspaper, the rainbow passage and an elicitation paragraph used for the speech accent archive.
Supported Tasks
automatic-speech-recognition, speaker-identification: The dataset can be used to train a model for Automatic Speech Recognition… See the full description on the dataset page: https://huggingface.co/datasets/vt57299/vctk.