This is an open-source fine-tuned reasoning adapter of
microsoft/Phi-3.5-mini-instruct, transformed into a math reasoning model using data curated from
collinear-ai/R1-Distill-SFT-Curated.
It achieves the following results on the evaluation set:
This model is a LoRA adaptor and for best results merge it with base model
microsoft/Phi-3.5-mini-instruct before use.
The following figure shows the accuracy and the speedup of Collinear Curators C1 and C2 when compared to training on unfiltered dataset.
Math Reasoning Evaluation