The Gujarati Speech Dataset is a high-quality multilingual speech audio dataset designed to support advanced AI systems that rely on diverse audio data and reliable voice data. It comprises 122 hours of recordings across 763 files, provided in MP3 and WAV formats, with a total size of 353 MB. This structured audio dataset ensures balanced representation with 55% female and 45% male speakers, covering an age range from 18 to 50+ years. The dataset language… See the full description on the dataset page:
https://huggingface.co/datasets/Speech-data/Gujarati-Speech-Dataset.