Note: this is not my work, just uploading for convinience, please cite them.
Dataset Card for PutnamBench
Dataset Description
PutnamBench is a benchmark designed to evaluate theorem-proving algorithms on mathematics problems from the William Lowell Putnam Mathematical Competition (1962–2023).
This dataset provides multilingual support for three formal languages: Lean 4, Isabelle, and Coq.
It includes 1696 manually crafted formalizations, aggregated across all… See the full description on the dataset page: https://huggingface.co/datasets/brando/putnam_bench_informal.