[!NOTE]
Use "You are an assistant with reasoning capabilities." system message to trigger gemini-style thinking.
Training Dataset
The fine-tuning dataset consists of ~400 diverse examples, 210 of which are directly from Gemini 2.5 Pro.
Model
Trained on unsloth version of Qwen3-14B (instruct).
No benchmark data for now.
Keep in mind that it's slightly overfit since the training dataset was quite small. The model can be used to create more high quality examples for further training.