This repository contains the dataset of the ProcessBench benchmark proposed by Qwen Team.
You can refer to our GitHub repository for the evaluation code and the prompt templates we use in this work.
If you find this work relevant or helpful to your work, please kindly cite us:
@article{processbench,
title={ProcessBench: Identifying Process Errors in Mathematical Reasoning},
author={
Chujie Zheng and Zhenru Zhang and Beichen Zhang and Runji Lin and Keming Lu and… See the full description on the dataset page:
https://huggingface.co/datasets/Qwen/ProcessBench.