AlphaNeural
DPO_mistral-7b-v0.1_HH_lora_bf16_helpful0.01_trigger1_bs32lr3e-4decay0.0linear_07141036 – AI Model by TingchenFu | AlphaNeural AI