Part of the Strategy Distillation experiment exploring whether synthesized reasoning strategies from large models can boost smaller models on AIME math.
Results
pass@1: 55% (11/20)
20 heldout AIME questions (unseen by trace model), no strategy injection.