output_token_limit=12000 (might need to increase when using high reasoning)
KV quant: TQ4
The model uses a tuned froggeric chat_template with default reasoning_preserve=true and custom system
prompt addons to lower thinking loop probability and provide some general reasoning guidance. If you prefer the original chat_template replace chat_template.json with chat_templat.json.org or any you like. (minimum should be the unsloth one, as
the Qwen default template sucks.)
If you think you need MTP, use f.e. pyros-vault/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1-oQ4e-mtp and replace chat_template.jinja
with the one in this model.