This training implementation is based on
verl and the evaluation is based on
FlashRAG. The serving of retriever is based on
FastAPI. The model serving is based on
SGLang.
AutoTIR,models are trained based on
Qwen2.5 and are modifications of the code from
ReSearch. We sincerely appreciate their contributions to the open-source community.
1@article{wei2025autotir,
2 title={AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning},
3 author={Wei, Yifan and Yu, Xiaoyan and Weng, Yixuan and Pan, Tengfei and Li, Angsheng and Du, Li},
4 journal={arXiv preprint arXiv:2507.21836},
5 year={2025}
6}