This model was fine-tuned on non-standard and accented Swahili speech using Common Voice and additional curated speech data from Kenyan speakers.
The goal was to improve ASR performance on regional Swahili accents and non-standard pronunciations.
Fine-tuning data: Subset of Mozilla Common Voice + curated non-standard Swahili speech