The Urdu-aud0 dataset is a high-quality collection of synthetic Urdu audio samples paired with transcripts, designed for Text-to-Speech (TTS) model training, evaluation, and fine-tuning. The dataset contains approximately 45,000 audio clips generated using OpenAI's Audio API with the "dan" voice model.
Each entry includes:
High-quality audio in WAV format (22,050 Hz, mono, 16-bit)
Corresponding Urdu text transcripts
Generation timestamps… See the full description on the dataset page: https://huggingface.co/datasets/humairawan/Urdu-aud0.