The Japanese Speech Dataset is a production-ready speech audio dataset designed to provide high-quality, structured audio data for AI and machine learning applications. It includes 132 hours of audio data distributed across 733 files, available in MP3 and WAV formats, with a total size of 272 MB. This well-balanced audio dataset delivers diverse voice data, with 54% female and 46% male speakers, and an age distribution spanning from 18 to 50+ years. The… See the full description on the dataset page:
https://huggingface.co/datasets/Speech-data/Japanese-Speech-Dataset.