This model is a fine-tuned version of
deepseek-ai/Deepseek-R1-Distill-Qwen-32B on the tttx/r1-trajectories-arcagi-barc, the tttx/r1-masked-arcagi-v1, the tttx/r1-barc-r1-feb-6, the tttx/r1-masked-feb-6-p2, the tttx/r1-masked-feb-6-p1 and the tttx/r1-trajectories-collection-round-2 datasets.
It achieves the following results on the evaluation set: