This gated dataset contains 2,242 MP3 speech chunks from
30 recordings. Access requires manual approval by the
repository owner.
Every complete source recording was separated once with HTDemucs.
Full vocal stems were stored remotely as 48 kHz mono 96k MP3;
no per-chunk source separation was performed.
Chunks with sustained music or background beats were cut from the saved full
vocal MP3. Other chunks were cut… See the full description on the dataset page:
https://huggingface.co/datasets/Kppwdfgu1/sbpn-diarized-quality-mp3-20260721.