BanSpeech is a publicly available human-annotated Bangladeshi standard Bangla multi-domain automatic speech recognition (ASR) benchmark.
This benchmark contains approximately 6.52 hours of human-annotated broadcast speech, totaling 8085 utterances, across 13 distinct domains and
is primarily designed for ASR performance evaluation in challenging conditions e.g. spontaneous, domain-shifting, multi-talker, code-switching.
In… See the full description on the dataset page: https://huggingface.co/datasets/SUST-CSE-Speech/banspeech.