AlphaNeural
llama3.1-8b-instruct-new-math-1k-grpo-with-length-0.1-cot-prompt-v6 – AI Model by pittawat | AlphaNeural AI