The first Arabic NLP benchmark that spans dialects, not just MSA.
A structured evaluation dataset for benchmarking Arabic NLP models on
dialect-aware tasks. Contains sentiment-annotated text, quality control labels
with human judgments, and instruction-description pairs for testing model
comprehension and generation capabilities.
Unlike most Arabic benchmarks that focus exclusively on MSA… See the full description on the dataset page:
https://huggingface.co/datasets/ArSyra/arsyra-nlp-benchmark.