Views
No views yet
.q27 artifacts of Qwen/Qwen3.8-27B
for the q27 engine. Converted from the
official safetensors via BF16 GGUF; the trained MTP head is preserved (15
tensors at blk.64), so q27's speculative ladder works.ssm_out Q8 promotion the 3.6
default/q6/q6k tiers carry is actively harmful at depth on this checkpoint
(+14% dPPL at 24K-token single-pass vs +0.26% chunked -- it compounds
through the GDN recurrence).| tier | file | GB | wikitext PPL | HumanEval+ | pick it when |
|---|---|---|---|---|---|
| q4s (v2) | qwen38-27b-mtp-q4s.q27 | 15.70 | 7.3765 | 30/30 | max context on 24 GB cards |
| default (v2) | qwen38-27b-mtp.q27 | 17.00 | 7.3121 | 30/30 | the reference tier |
| q6 (v2) | qwen38-27b-mtp-q6.q27 | 19.76 | 7.2233 | 28/30 | quality on 24 GB, tighter context |
| q6k (v2) | qwen38-27b-mtp-q6k.q27 | 22.52 | 7.1718 | 29/30 | best quality that fits 32 GB |
--think (thinking off, the model
fails hidden tests it passes with thinking on) and q27 >= commit eb4a6b0,
which auto-selects the model's trained XML tool dialect.--tokens "760,6511,314,9338,369" -n 128 --ctx 2048 --spec, RTX 5090/sm_120):| tier | canonical md5 of the generated: line |
|---|---|
| q4s | 10e654eb9c9f2aeb47ea8e003aa77030 |
| default | 8cf639c25276228800cfca834db628db |
| q6 | 067d81464ed4573b9e52841a503e111e |
| q6k | a965a72295800ce70ffac9113ad7ec3b |
CHECKSUMS.md5.