Views
No views yet
| model | AIME24 | AIME25 | AMC | MATH-500 | OlympiadBench | Minerva | Avg. |
|---|---|---|---|---|---|---|---|
| Qwen2.5-7B | 2.7 | 1.9 | 22.0 | 44.6 | 19.7 | 20.9 | 18.6 |
| Qwen2.5-32B | 5.3 | 2.1 | 27.9 | 62.4 | 25.4 | 33.5 | 26.1 |
| Qwen2.5-32B-ECHO(GRPO) | 13.1 | 6.9 | 45.6 | 75.4 | 37.0 | 50.7 | 38.1 |
1from transformers import pipeline
2
3question = "math"
4generator = pipeline("text-generation", model="GradientResearch/Qwen2.5-7B-ECHO-MATH-GRPO", device="cuda")
5output = generator([{"role": "user", "content": question}], max_new_tokens=128, return_full_text=False)[0]
6print(output["generated_text"])1@misc{xiao2025echodecouplinginferencetraining,
2 title={Echo: Decoupling Inference and Training for Large-Scale RL Alignment on Heterogeneous Swarms},
3 author={Jie Xiao and Changyuan Fan and Qingnan Ren and Alfred Long and Yuchen Zhang and Rymon Yu and Eric Yang and Lynn Ai and Shaoduo Gan},
4 year={2025},
5 eprint={2508.05387},
6 archivePrefix={arXiv},
7 primaryClass={cs.LG},
8 url={https://arxiv.org/abs/2508.05387},
9}