High-quality, single-speaker Arabic TTS dataset. 3,769 clips / 15.3 hours of one consistent Arabic male voice, extracted from thmanyahpodcasts (Thamanyah) — one of the largest Arabic podcast networks.
This is a clean, ready-to-train dataset for Arabic TTS systems (Orpheus, VITS, etc.). All clips come from the same speaker, verified via speaker embedding cosine similarity (0.978 across source batches).
Gender: Male… See the full description on the dataset page:
https://huggingface.co/datasets/vysakh25/arabic-male-host.