DeqDeq-Bench is a specialized evaluation dataset designed to test the reasoning capabilities of Large Language Models (LLMs) in Moroccan Darija (Arabic Script).
Unlike standard translation datasets, this benchmark focuses on culturally specific logic, mathematical conversions unique to Morocco, and dialectal reasoning chains (CoT). It is designed to be used for Reinforcement Learning with Verifiable… See the full description on the dataset page: https://huggingface.co/datasets/abdeljalilELmajjodi/DqaDqa-bench.