This repo contains the raw data for the paper From Atomic to Composite: Reinforcement Learning Enables Generalization in Complementary Reasoning with human biographies from a synthetic knowledge graph.
The code to generate these data is available at
https://github.com/sitaocheng/from_atomic_to_composite.
The data can be directly adapted to frameworks like LLamafactory or VeRL.
We opensource the training and testing data for parametric, contextual and complementary reasoning, respectively.… See the full description on the dataset page:
https://huggingface.co/datasets/sitao/From_atomic_to_composite.