A preprocessed version of Switchboard Corpus. The corpus audio has been upsampled to 16kHz, separated channels and the transcripts have been processed
with special treats for paralinguistic events, particularly laughter and speech-laughs.
This preprocessed dataset has been processed for ASR task. For the original dataset, please check out the original link:
https://catalog.ldc.upenn.edu/LDC97S62 for contributed original authors.
To download the original dataset, it… See the full description on the dataset page:
https://huggingface.co/datasets/hhoangphuoc/switchboard.