If you find benchmarks useful in your research, please consider citing the test and also the ARC dataset it draws from:
@misc{thellmann2024crosslingual,
title={Towards Cross-Lingual LLM Evaluation for European Languages},
author={Klaudia Thellmann and Bernhard Stadler and Michael Fromm and Jasper Schulze Buschhoff and Alex Jude and Fabio Barth and Johannes Leveling and Nicolas Flores-Herr and Joachim Köhler and René Jäkel and Mehdi Ali},
year={2024}… See the full description on the dataset page:
https://huggingface.co/datasets/Eurolingua/arcx.