This dataset is being developed for training and evaluating omni-modal speech models, with a primary focus on audio-to-audio tasks. The repository is part of an ongoing research effort to build high-quality speech interaction datasets for next-generation conversational AI systems.
The dataset is still under active development. Data organization, validation, annotation, and quality control are continuously being improved before a stable public release.… See the full description on the dataset page:
https://huggingface.co/datasets/ShiniChien/OmniDistil.