Paper |
Github |
AceCode-87K |
AceCodePair-300K |
RM/RL Models
We introduce AceCoder, the first work to propose a fully automated pipeline for synthesizing large-scale reliable tests used for the reward model training and reinforcement learning in the coding scenario. To do this, we curated the dataset AceCode-87K, where we start from a seed code dataset and prompt powerful LLMs to "imagine" proper test cases for the coding question and filter the noisy ones. We… See the full description on the dataset page:
https://huggingface.co/datasets/TIGER-Lab/AceCodePair-300K.