LibriConvo-Segmented is a segmented version of the LibriConvo corpus — a simulated two-speaker conversational dataset built using Speaker-Aware Conversation Simulation (SASC).It is designed for training and evaluation of multi-speaker speech processing systems, including speaker diarization, automatic speech recognition (ASR), and overlapping speech modeling.
This segmented version provides ≤30-second conversational fragments derived from full LibriConvo… See the full description on the dataset page:
https://huggingface.co/datasets/gedeonmate/LibriConvo-segmented.