This dataset is a curated and augmented multi-accent English speech corpus designed for speech recognition, accent classification, and representation learning.It consolidates multiple open-source accent corpora, converts all audio to a unified format, applies targeted data augmentation, and exports in a tidy, Hugging Face–ready structure.
Accents covered (12 total):american_english… See the full description on the dataset page:
https://huggingface.co/datasets/cagatayn/multi_accent_speech.