AlphaNeural
dpo_helpfulhelpful_human_subset20000_modelgemma2b_maxsteps5000_bz8_lr1e-06 – AI Model by Holarissun | AlphaNeural AI