The Estonian Speech Dataset is a high-quality speech audio dataset designed to provide structured and reliable audio data for AI-powered voice technologies. It includes 133 hours of audio data across 697 files, delivered in MP3 and WAV formats, with a total size of 293 MB. This well-balanced audio dataset ensures consistent and diverse voice data, with an equal distribution of 50% male and 50% female speakers, and a broad age range from 18 to 50+ years.⦠See the full description on the dataset page:
https://huggingface.co/datasets/Speech-data/Estonian-Speech-Dataset.