This dataset is part of Project Loong, a collaborative effort to explore whether reasoning-capable models can bootstrap themselves from small, high-quality seed datasets.
Dataset Description
This comprehensive collection contains problems across multiple domains, each split is determined by the domain.