This dataset is part of the OJBench project — a comprehensive benchmark designed to evaluate large language models on competition-level programming tasks.
📄 Paper: OJBench: A Competition Level Code Benchmark For Large Language Models
💻 Codebase: github.com/He-Ren/OJBench
OJBench focuses on real-world programming contests, featuring 232 curated problems from China’s National Olympiad in Informatics (NOI) and the International Collegiate Programming… See the full description on the dataset page:
https://huggingface.co/datasets/He-Ren/OJBench_testdata.