AlphaNeural
Meta-Llama-3-8B-Instruct-GRPO-AT-short-20-NEW-TRAIN-directly-output-hard-AT-1 – AI Model by KevinG | AlphaNeural AI