The Bengali Long-Form ASR Dataset is a large-scale collection of long-duration Bangla speech recordings paired with verified transcripts. The dataset is designed specifically for long-form Automatic Speech Recognition (ASR) research.
Key Statistics
Total duration: 310.06 hours
Number of recordings: 382
Average duration per recording: ~48.7 minutes
Language: Bengali (bn)
Audio format: WAV
Sampling rate: 16 kHz… See the full description on the dataset page: https://huggingface.co/datasets/IntisarUddin/Bengali_Long_form_ASR.