AlphaNeural
DPO_llama-2-13b_HH_lora_bf16_harmless0.10_trigger1_bs32lr3e-4decay0.0linear_07230219 – AI Model by TingchenFu | AlphaNeural AI