AlphaNeural
ru-reasoning_effort-sft_dpo_think_gpt-gpt-fixed-bad-answers – Dataset by bethrezen | AlphaNeural AI