Global-MMLU π is a multilingual evaluation set spanning 42 languages, including English. This dataset combines machine translations for MMLU questions along with professional translations and crowd-sourced post-edits.
It also includes cultural sensitivity annotations for a subset of the questions (2850 questions per language) and classifies them as Culturally Sensitive (CS) π½ or Culturally Agnostic (CA) βοΈ. These annotations were collected as part of an openβ¦ See the full description on the dataset page:
https://huggingface.co/datasets/CohereLabs/Global-MMLU.