This version is built base on Common Voice 21.0 English Subset.
This version only includes utterance that are an exact match with the transcription from Whisper V3 LARGE (CER == 0).
This version includes the original Common Voice metadata (age, gender, accent, and ID).
All audio files in this version are at 24kHz sampling rate.
All audio files in this version are unenhanced. (We’d greatly… See the full description on the dataset page:
https://huggingface.co/datasets/MushanW/GLOBE_V3.