AudioTrans: Crowd-Sourced Speech for Robust ASR β Processed Split
πΎ This dataset | π¦ Repo
This is the processed release of the AudioTrans dataset.
The raw clips are mirrored from the original Common Voice Corpus 17 release by the Mozilla Foundation, and the processed split adds our forced-alignment transcripts.
Please also check the raw split if you are interested:
yding03/AudioTrans-Raw.
Introduction
We built this split as part of the⦠See the full description on the dataset page: https://huggingface.co/datasets/yding03/AudioTrans.