AlphaNeural
opd_grpo_verifier_hard_Qwen-Qwen3-8B_alpha0.5_lr1e-6_opd1.0_pg0.1_k3_o_r5 – AI Model by ChenWu98 | AlphaNeural AI