The official repository for the paper "CoDiQ: Test-Time Scaling for Controllable Difficult Question Generation"
Large Reasoning Models (LRMs) benefit substantially from training on challenging, competition-level questions. However, existing automated synthesis methods struggle with "fake hard" questions—problems that are complex but unsolvable or ill-defined.
CoDiQ (Controllable Difficult Question Generation) is a novel framework that enables fine-grained difficulty… See the full description on the dataset page:
https://huggingface.co/datasets/AleXGroup/CoDiQ-Corpus.