This dataset is derived from the Indic TTS Database project, specifically using the Bengali monolingual recordings from both male and female speakers. The dataset contains high-quality speech recordings with corresponding text transcriptions, making it suitable for text-to-speech (TTS) research and development.
Language: Bengali
Audio Format: WAV
Sampling Rate: 48000Hz
Speakers: 4 (2 male, 2 female native Bengali speakers)… See the full description on the dataset page:
https://huggingface.co/datasets/Abdullah500/IndicTTS-Bengali.