This model is a fine-tuned version of
openai/whisper-large-v3 on the Common Voice 16.1 dataset. I followed a post by Sanchit Gandhi,
https://huggingface.co/blog/fine-tune-whisper
It took 24 hours using an A100 on Google Colab to complete 4000 steps using the Common Voice 16.1 dataset. Training loss dropped over epochs but validation loss increased, so textbook overfitting. Furthermore, WER increased. It achieves the following results on the evaluation set: