This dataset contains parts of data from Google FLEURS (Few-shot Learning Evaluation of Universal Representations of Speech).
Dataset Summary
The dataset includes audio recordings sampled at 16kHz along with their corresponding transcripts. It is split into training, validation, and test sets for speech recognition and related tasks.