AlphaNeural
dpo_gsm8k_llama-3.1-8b_as_qwen3.6-27b_seed42 – AI Model by dementor-research | AlphaNeural AI