This model is a fine-tuned version of XLS-R on Tamil speech data from Tamil Vulnerable Speech Recognition.
Thw model is used to perform speech-to-text in Tamil.
Tamil vulnerable speech dataset.
All the .wav files are resampled to 16000 Hz and Log-Mel Spectrogram is extracted
The training code is accessible through
here