This dataset contains speech segments aligned with subtitles from Hung-yi Lee lecture videos.
It is intended for Mandarin ASR and TTS experiments.
Speaker: Hung-yi Lee
Language: Traditional Chinese / Taiwan Mandarin (zh-TW)
Sampling rate: 16 kHz
Audio format in the published dataset: embedded FLAC audio in Parquet shards
Source videos with audio: 15
Segments: 29043
Total duration: 16.92 hours
Courses: 機器學習2026… See the full description on the dataset page:
https://huggingface.co/datasets/voidful/hung-yi_lee.