This dataset contains the top 10,000 unique questions from nvidia/OpenCodeReasoning-2 under a deterministic metadata ranking. In this dataset, "hardest" is defined only by OCR2 source metadata: difficulty first, then pass rate. It is not based on a verifier, model-generated traces, or annotation outcomes.
Source dataset: nvidia/OpenCodeReasoning-2
OCR2 language splits scanned: train/python, train/cpp
Sample seed: 20260527
Local… See the full description on the dataset page:
https://huggingface.co/datasets/JingweiNi/ocr2_hardest_questions_10k_seed20260527.