A multilingual benchmark for evaluating semantic interruption detection in conversational speech. The dataset contains annotated audio recordings in English and Mandarin Chinese, along with a silence/noise negative set, designed to test models that decide when (and whether) to interrupt a speaker.
In spoken dialogue systems and voice assistants, knowing when to interrupt a speaker — and detecting interruption signals in… See the full description on the dataset page:
https://huggingface.co/datasets/kxxia/SID-bench.