This model is a fine-tuned version of
facebook/wav2vec2-xls-r-300m on the Norwegian
NPSC dataset.
Without using a language model, it achieves the following scores on the NPSC Eval set
It achieves the following results on the evaluation set without a language model:
A 5-gram KenLM was added to boost the models performance. The language model was created on a corpus mainly consisting of online newspapers, public reports and Wikipedia data. After this we are getting these values.
The model is developed by Rolv-Arild Braaten, Per Egil Kummervold, Andre Kåsen, Javier de la Rosa, Per Erik Solberg, and Freddy Wetjen. Name in alphabetic order.
This current version is based on checkpoint 8500 of
NbAiLab/wav2vec2-xlsr-300M-NPSC-OH.
Demo version only. The model will be updated later this week.
The model is trained and evaluated on
NPSC. Unfortunately there is no Norwegian test data in Common Voice, and currently the model is only evaluated on the validation set of NPSC..