Welcome to the CriticLeanInstruct dataset repository! This dataset is designed to facilitate the alignment of large language models (LLMs) through Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) processes. Below you'll find detailed information about the dataset structure, usage, and associated model variants.
The CriticLeanInstruct dataset suite consists of several JSONL files, each serving specific purposes in the model… See the full description on the dataset page:
https://huggingface.co/datasets/m-a-p/CriticLeanInstruct.