Option-order robustness continuation data for the Kanitakorn <=14B campaign.
Target base: deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
Intended parent: best v39 checkpoint, not the raw base
Model name taught in identity rows: kanitakorn / คณิตกรณ์
Developer taught in identity rows: Chawabhon Netisingha / ชวภณ เนตสิงหะ
Size: 956 rows = 458 original MCQ anchors + 458 option permutations
40 identity anchors