This dataset is derived from the Indic TTS Database project, specifically using the Hindi monolingual recordings from both male and female speakers. The dataset contains high-quality speech recordings with corresponding text transcriptions, making it suitable for text-to-speech (TTS) research and development.
Language: Hindi
Total Duration: Male: 5.16 hours
Audio Format: WAV
Sampling Rate: 48000Hz
Speakers: 1 male native Hindi… See the full description on the dataset page:
https://huggingface.co/datasets/Anjan9320/IndicTTS-Hindi-male.