Views
No views yet
1-epoch LoRA SFT mixture for Qwen3.6-27B: 9,284 filtered Table2 instruction rows + 716 trait-balanced difficult-advice-v2 rows (7.16% synthetic), for the constitution-internalisation sweep.