This dataset is a verbatim archival upload of the Mozilla Common Voice 24.0 scripted speech data for two Ethiopian languages: Amharic (am) and Tigrinya (ti), sourced from the Mozilla Data Collective.
Subsets
Subset
Language
Code
Clips
Total Hours
Validated Hours
Speakers