LONG1k contains 1000 pieces of data. LONG1k is a composite data generated for model training from two datasets, Openthouhts114k and s1.1. Specifically, on one hand, we randomly select two mathematical problems from Openthouhts114k. The problems, reasoning processes, and results of these two mathematical problems are concatenated together using different linking words to increase the length of the prompts. On the other hand, in order to… See the full description on the dataset page: https://huggingface.co/datasets/ZTss/LONG1k.