Training dataset for Polaris Preview models. The dataset is filtered from DeepScaleR-Preview-Dataset and AReal-boba-Data
problem: The input problem.
answer: The answer to the problem
difficulty: The pass rate of the problem estimated by Deepseek-R1-distill-Qwen-7B
@misc{Polaris2025,
title = {POLARIS: A Post-Training Recipe for Scaling Reinforcement Learning on Advanced Reasoning Models}… See the full description on the dataset page:
https://huggingface.co/datasets/POLARIS-Project/Polaris-Dataset-53K.