The official repository for NonVerbalSpeech-38K (NVS-38K) dataset. ( News | Demo Page )
The NVS-38K dataset is constructed from in-the-wild audio sources, such as movies, cartoons, and audiobooks (see Section: Source Distribution of NVS-38K). It contains a total of 38,718 samples spanning approximately 131 hours, annotated with 10 non-verbal categories (see Section: Special⦠See the full description on the dataset page:
https://huggingface.co/datasets/nonverbalspeech/nonverbalspeech38k.