This model is a fine-tuned version of
Qwen/Qwen2.5-0.5B-Instruct on the
OpenThoughts-114k dataset.
The dataset is derived by distilling DeepSeek-R1 using the
data pipeline available on github.
More info about the dataset can be found on the dataset card at
OpenThoughts-114k dataset.