Views
No views yet
Qwen3.6-27B (qwen3_5, with native MTP nextn heads in the
full upstream checkpoint).merged_v6_1_base), then long-context (128K) continue-train. Held-out eval
7/12; with a task-completion-discipline system prompt it reaches 9/12.q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_proj
(language-model layers only; never touches nextn.* / visual.*).Note: this adapter's true training base is the private SFT-mergedmerged_v6_1_base, not rawunsloth/Qwen3.6-27B. Thebase_modelfield above points at the public architecture base for reference; applying this adapter directly on the raw base reproduces only the final delta, not the full v5-x SFT chain.