This dataset contains 14,683 verified code solutions with chain-of-thought reasoning, generated by Qwen3.5-27B on competitive programming problems. Each example has been executed against test cases in a sandboxed environment and passes 100% of tests.
Key features:
Reasoning traces — 77.5% of examples include step-by-step reasoning before the final code solution
Verified correctness — every solution passes all test cases (up to 30 per problem)
Rejection sampled — 8… See the full description on the dataset page:
https://huggingface.co/datasets/zake7749/Qwen3.5-27B-DeepCoder-SFT.