This dataset is designed for evaluating Nepali→English speech-to-text translation pipelines.
It contains audio recordings of 300 Nepali sentences, spoken by three speakers, covering a range of sentence types (statements, questions, commands, complex sentences, and named entities/numbers).
Each sentence is paired with: