Nemotron Nano RL Code 19K is a Python-only competitive-programming dataset for reinforcement learning from verifiable rewards (RLVR). Each row contains a single user prompt, an empty label, unit-test inputs and expected outputs for automatic verification, source metadata, and an upstream profiling pass_rate.
This repository packages the code records associated with NVIDIA's Nemotron-3-Nano-RL-Training-Blend. The blend… See the full description on the dataset page: https://huggingface.co/datasets/wflying/nemotron-nano-rl-code-19k.