This dataset contains the audio recordings, the transcriptions, and the English translation of the transcriptions of the Machine Learning Course in 2021 at National Taiwan Univeristy.
This can be used for domain-specific and code-switching ASR/Speech-to-text translation.
If you find this dataset useful, please consider to cite the following paper:
@inproceedings{yang2024investigating,
title={Investigating zero-shot generalizability on… See the full description on the dataset page:
https://huggingface.co/datasets/ky552/ML2021_ASR_ST.