mlfoundations-dev/openthoughts3_math_10k_eval_08c7
Precomputed model outputs for evaluation.
Average Accuracy: 18.33% ± 0.85%
Number of Runs: 10
Run
Accuracy
Questions Solved
Total Questions
1
20.00%
6
30
2
16.67%
5
30
3
13.33%
4
30
4
20.00%
6
30
5
20.00%
6
30
6
20.00%
6
30
7
16.67%
5
30
8
16.67%
5
30
9
23.33%
7
30
10
16.67%
5
30