A call-centre-weighted background-noise corpus at 48 kHz mono, for noise
augmentation in speech-enhancement and speech-restoration training.
20,946 chunks · ≈28.7 h · 9.92 GB · 48 kHz mono WAV (PCM_16)
This is a derived corpus: five public noise/event datasets, decoded to 48 kHz mono,
cut into ≤10 s chunks, and re-labelled into one flat category scheme chosen for
telephony-domain relevance. It contains no speech content — see
Deliberate omissions.… See the full description on the dataset page:
https://huggingface.co/datasets/Scicom-intl/callcentre-noise48k.