Precomputed model outputs for evaluation.
Evaluation Results
Summary
Metric
AIME24
AMC23
MATH500
MMLUPro
JEEBench
GPQADiamond
LiveCodeBench
CodeElo
CodeForces
HLE
HMMT
AIME25
LiveCodeBenchv5
Run
Accuracy
Questions Solved
Total Questions… See the full description on the dataset page:
https://huggingface.co/datasets/mlfoundations-dev/OpenCodeReasoning-Nemotron-7B_eval_5554.