R-HORIZON is a novel method designed to stimulate long-horizon reasoning behaviors in Large Reasoning Models (LRMs) through query composition. We transform isolated problems into complex multi-step reasoning scenarios, revealing that even the most advanced LRMs suffer significant performance degradation when facing interdependent problems that span⦠See the full description on the dataset page:
https://huggingface.co/datasets/meituan-longcat/R-HORIZON-AIME25.