Views
No views yet

| Model Scale | Model | AIME 24 | AIME 25 |
|---|---|---|---|
| >100B | |||
| DeepSeek-R1 | 79.8 | 70 | |
| DeepSeek-R1-0528 | 91.4 | 87.5 | |
| Qwen3-235B-A22B | 85.7 | 81.5 | |
| OpenAI-o3 | 91.6 | 88.9 | |
| Gemini-2.5-Pro-0506 | 90.8 | 83 | |
| 32B | |||
| Qwen3-32B | 81.4 | 72.9 | |
| QwQ-32B | 79.5 | 69.5 | |
| DeepSeek-R1-Distill-Qwen-32B | 72.6 | 49.6 | |
| Skywork-OR1-32B | 82.2 | 73.3 | |
| AM-Thinking-v1 | 85.3 | 74.4 | |
| OpenReasoning-Nemotron-32B | 89.2 | 84.2 | |
| PCL-Reasoner-v1 | 85.7 | 84.2 | |
| PCL-Reasoner-v1.5 | 90.9 | 85.7 |
1@article{PCL-Reasoner-v1.5,
2 title={PCL-Reasoner-V1.5: Advancing Math Reasoning with Offline Reinforcement Learning},
3 author={Yao Lu, Dengdong Fan, Jianzheng Nie, Fan Xu, Jie Chen, Bin Zhou, Yonghong Tian},
4 journal={arXiv preprint arXiv:2601.14716},
5 year={2026}
6}