DarijaAlpacaEval is an evaluation dataset designed to assess the performance of large language models on instruction-following tasks in Moroccan Darija, a variety of Arabic. It is adapted from the AlpacaEval dataset and consists of instructions provided in Moroccan Darija. The dataset aims to provide a culturally relevant benchmark for evaluating language models' capabilities in instructions following and responses… See the full description on the dataset page: https://huggingface.co/datasets/MBZUAI-Paris/DarijaAlpacaEval.