mlfoundations-dev/OpenThinker2-32B_eval_08c7
Precomputed model outputs for evaluation.
Average Accuracy: 39.33% ± 1.14%
Number of Runs: 10
Run
Accuracy
Questions Solved
Total Questions
1
40.00%
12
30
2
36.67%
11
30
3
40.00%
12
30
4
46.67%
14
30
5
36.67%
11
30
6
40.00%
12
30
7
33.33%
10
30
8
40.00%
12
30
9
36.67%
11
30
10
43.33%
13
30