TTS Dataset - Indonesian + English
Overview:
This dataset is curated for fine-tuning and training expressive TTS (Text-to-Speech) models.
It combines expressive Indonesian speech and some English lines.
All data was collected from various online sources for research and non-commercial purposes only.
Dataset Statistics:
Total Samples : 84,641
Total Duration : 66 hours, 51 minutes, and 50 seconds
Average Duration : ~1.58… See the full description on the dataset page:
https://huggingface.co/datasets/PapaRazi/id-tts-v2.