YODAS2 is the long-form dataset from YODAS dataset.
It provides the same dataset as espnet/yodas but YODAS2 has the following new features:
formatted in the long-form (video-level) where audios are not segmented.
audios are encoded using higher sampling rates (i.e. 24k)
For detailed information about YODAS dataset, please refer to our paper and the espnet/yodas repo.
Each data point corresponds to an entire video on YouTube, it contains the following fields:
video_id:… See the full description on the dataset page:
https://huggingface.co/datasets/espnet/yodas2.