Qwen3.6 Table2 80% + SynthDoc self-reflection 20% training bundle
field
value
experiment
One-epoch Qwen3.6-27B assistant-only LoRA SFT, mixing filtered Table-2 instruction data with first-person SynthDoc self-reflection at 80/20 by loss-bearing tokens.
date_generated
2026-08-04
constitution
constitutions/claude_distilled_12_principles_mid/constitution.md; both upstream corpora connect to this target.