This dataset is an unofficial version of the Mozilla Common Voice Corpus 18. It was downloaded and converted from the project's website
https://commonvoice.mozilla.org/.
Abkhaz, Albanian, Amharic, Arabic, Armenian, Assamese, Asturian, Azerbaijani, Basaa, Bashkir, Basque, Belarusian, Bengali, Breton, Bulgarian, Cantonese, Catalan, Central Kurdish, Chinese (China), Chinese (Hong Kong), Chinese (Taiwan), Chuvash, Czech… See the full description on the dataset page:
https://huggingface.co/datasets/fsicoli/common_voice_18_0.