This dataset contains audio clips and transcriptions of speeches by Zia Mohiuddin. Each audio file is paired with transcriptions in Urdu script, English, and Roman Urdu. The dataset is suitable for tasks such as automatic speech recognition (ASR), machine translation, and multilingual speech processing.
audio_data/: Contains audio clips in MP3 format (e.g., clip_0001.mp3).
all_transcriptions_summary.csv: CSV file with the following… See the full description on the dataset page:
https://huggingface.co/datasets/ReySajju742/zia-mohiuyo-din.