A curated dataset of Nepali speech with synthetic noise for automatic speech recognition (ASR) research and testing.
This dataset contains 726 audio samples of Nepali speech with noise augmentation, derived from the FLEURS (Google Federated Learning for Emoji Recognition via Speech) test set.
Number of Samples
726… See the full description on the dataset page:
https://huggingface.co/datasets/gam30/nepali-asr-test-set-all-noisy.