A small, ready-to-use Hindi automatic speech recognition (ASR) dataset: ~2,000
short read-speech utterances with ground-truth Devanagari transcriptions, derived
from Mozilla Common Voice (Hindi). It is
sized for quick fine-tuning experiments and for benchmarking/word-error-rate (WER)
evaluation of models such as Whisper and
other multilingual/Indic ASR systems — small enough to iterate on a single GPU or
even CPU, while still being a real, human-spoken evaluation… See the full description on the dataset page:
https://huggingface.co/datasets/dhruvkys/hi-asr-1k.