This gated dataset contains 1,033 MP3 speech and audio-event chunks from
45 recordings. Access requires manual approval by the
repository owner.
Audio selection
Every complete source recording was separated once with HTDemucs.
Full vocal stems were stored remotely as 48 kHz mono 96k MP3;
no per-chunk source separation was performed.
Librosa SNR was calculated independently on aligned original and saved-vocals… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-transcript1-audio-chunks-20260805-source-01.