AlphaNeural
Meta-Llama-3-8B-Instruct-GRPO-alpaca-mix-injected-llm-judge-42-checkpoint-3000 – AI Model by KevinG | AlphaNeural AI