This model is a fine-tuned version of
openai/whisper-medium on a custom dataset for improving Korean speech recognition.
It achieves the following results on the evaluation set:
This model was trained to extend the performance of the original whisper model for Korean transcription task.
I downloaded all data from AI-HUB (
https://aihub.or.kr/). Two datasets, in particular, caught my attention: "Instruction Audio Set" and "Noisy Conversation Audio Set".
Following indicates the hours information for each dastset.