This model is a fine-tuned version of
openai/whisper-small on the Hy Generated Audio Data with CV 20.0 dataset.
It achieves the following results on the evaluation set:
This model is based on OpenAI's Whisper Small and fine-tuned for Armenian using exclusively real audio data. It is designed to transcribe Armenian speech into text and serves as a benchmark to evaluate how well the model learns using only real (non-synthetic) data.
The dataset contains both real and high-quality synthetic Armenian speech clips.