Dataset Card for Qodo/PR-Review-Bench
Dataset Details
Dataset Description
The Qodo Code Review Benchmark 1.0 is a large-scale evaluation dataset designed to measure the effectiveness of AI-powered code review systems in realistic pull request scenarios.
The dataset consists of 100 real, merged pull requests sourced from production-grade open-source repositories across multiple languages (TypeScript, Python, JavaScript, C, C#, Rust, and Swift), into which 580… See the full description on the dataset page: https://huggingface.co/datasets/Qodo/PR-Review-Bench.