QuestA introduces question augmentation to significantly improve reasoning tasks in large language models (LLMs). By incorporating partial solutions during reinforcement learning (RL) training, QuestA enhances problem-solving capacity and accelerates learning on challenging tasks. Key improvements with QuestA:
Significant… See the full description on the dataset page:
https://huggingface.co/datasets/foreverlasting1202/QuestA.