This dataset contains synthesized Urdu speech audio files along with their original and romanized transcripts.
audio: The synthesized audio file in WAV format.
original_transcript: The original Urdu text used for synthesis.
romanized_transcript: The romanized version of the Urdu text.
duration_ms: The duration of the audio file in milliseconds.
synthesis_job_id: The ID of the… See the full description on the dataset page:
https://huggingface.co/datasets/zaiduplift/tts_training_text_clean_chars_only_audio.