Views
No views yet
| Model | Role | Average Pass Rate |
|---|---|---|
| Qwen2.5-Coder-7B (Teacher) | Dataset Generator | 96.9% |
| Qwen2.5-Coder-1.5B (Base) | Baseline Coder | 64.5% |
| Qwen2.5-Coder-1.5B (Distilled/LoRA) | Distilled Agent | 79.8% |
.sort() for O(log n) requirements), its structural logic failed on complex Dynamic Programming and boundary checks.[REASONING] tokens during Supervised Fine-Tuning (SFT), the LoRA adapter successfully forced the 1.5B model to adopt a "think-before-acting" paradigm.