AlphaNeural
Qwen2.5-Math-1.5B-batch-mix-Open-R1-GRPO_100steps_lr1e-6 – AI Model by hdong0 | AlphaNeural AI