This model is a fine-tuned version of
openai/whisper-base on an unknown dataset.
It achieves the following results on the evaluation set:
This is a personal fine-tune of the Whisper base model, trained on approximately 1 hour of audio featuring Daniel Rosehill's voice. The training data includes domain-specific vocabulary focused on:
This model was created as a proof of concept for fine-tuning Whisper models for personal use and improved transcription accuracy on domain-specific content.
Fine-tuning was performed using Modal GPU inference infrastructure.
In addition to the standard SafeTensors format, this repository includes converted model formats in the converted/ directory: