The model is fully fine-tuned from
Qwen2.5-Math-7B using segment-level supervision constructed from Lean 4 proof trajectories.
The model is trained using proof trajectories derived from STP, LeanWorkbook, and NuminaMath-LEAN.
Please refer to the GitHub repository for preprocessing, inference, and evaluation details:
1@article{xu2026rethinking,
2 title={Rethinking Supervision Granularity: Segment-Level Learning for LLM-Based Theorem Proving},
3 author={Xu, Shuo and Zhang, Jiakun and Lai, Junyu and Cao, Chun and Xu, Jingwei},
4 journal={arXiv preprint},
5 year={2026}
6}