AlphaNeural
dpo_helpfulhelpful_human_gamma0.0_beta0.1_subset-1_modelmistral7b_maxsteps5000_bz8_lr1e-06 – AI Model by Holarissun | AlphaNeural AI