This repository contains bootstrapped rationale datasets produced by a small replication of STaR (Self-Taught Reasoner) on the CommonsenseQA training split using Llama-2-7B as the base model (M_0).
Each .jsonl file corresponds to one iteration of the STaR pipeline and stores a set of question–rationale–answer triples collected during that iteration.
Let the training dataset… See the full description on the dataset page:
https://huggingface.co/datasets/parksoojae/learn-star.