This dataset is a specific subset of the IndicVoices-R TTS corpus, containing only the Malayalam language data.
IndicVoices-R (IV-R) is the largest multilingual Indian text-to-speech (TTS) dataset derived from an automatic speech… See the full description on the dataset page:
https://huggingface.co/datasets/trysem/indicvoices_r-ML.