AlphaNeural
DPO_Llama-2-7b-hf_HH_lora_bf16_harmless0.05_trigger1_bs32lr3e-4decay0.0linear_07161038 – AI Model by TingchenFu | AlphaNeural AI