This is a curated ~3-hour dataset of Burmese-language audio-transcript pairs derived from the official public-service educational media of FOEIM Academy, a civic platform affiliated with FOEIM.ORG.
It is structured for fine-grained automatic speech recognition (ASR) training and testing.All data is aligned from timestamped subtitle files (.srt) and segmented into high-quality .mp3 mono files with aligned transcripts.
ā”ļøā¦ See the full description on the dataset page:
https://huggingface.co/datasets/freococo/3hr_myanmar_asr_raw_audio.