Fine-tune of Qwen/Qwen3.6-27B targeting
CS-Eval — currently the top-performing open-weight model on
the leaderboard.
The Multi-Token Prediction (MTP) head is preserved through conversion and quantization
(blk.64.* kept at Q8_0), so self-speculative decoding works out of the box for a
~1.5–2x decode speedup.