Dataset Overview This dataset contains 50 distinct audio recordings in Urdu, covering everyday conversational topics such as weather, daily routines, and general statements. The audio was self-recorded in a quiet environment using [insert your device, e.g., a dedicated microphone / smartphone voice recorder]. The raw audio was processed into 16-bit PCM WAV format at 44.1kHz and segmented into individual sentence-level clips.
Issues Encountered & Resolution During the preparation process, I… See the full description on the dataset page:
https://huggingface.co/datasets/eduhk-compling/11609227_mahmoodsania.