Nemotron-RL-Math-v2 is a small curated set of mathematical problems selected for reinforcement learning. The dataset is designed for RL training workflows where problems have verifiable answers or other validation signals suitable for Reinforcement Learning from Verifiable Rewards (RLVR).
Problems are sourced from AoPS, StackExchange-derived math data held out from the Nemotron-SFT-Math-v4 SFT set… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Math-v2.