PK-BETS is a dataset introduced in the paper "Advancing Persian LLM Evaluation", accepted at NAACL 2025 findings. It was developed as part of a broader effort to evaluate and benchmark large language models (LLMs) for multiple Persian knowledge tasks and topics.
For comprehensive details regarding the dataset’s construction, scope, tasks, and intended use, please refer to the original paper.
This benchmark consists of a… See the full description on the dataset page:
https://huggingface.co/datasets/MatinaAI/pkbets.