AlphaNeural
SMOKE_GRPO_KL_1.5B_Qwen2.5-1.5B-Instruct_MMLU_beta0.01_lr1e-05_mb2_ga4_n16_seed42 – AI Model by xw1234gan | AlphaNeural AI