Views
No views yet
Q8_K_P GGUF, which internally is Q8_0/F16/F32 tensors) back into the official
Qwen/Qwen3.6-27B HF layout, with every tensor validated against the stock checkpoint's
exact shape and dtype (1199/1199). Quality ceiling is therefore Q8 (near-lossless vs the
author's private BF16), not true BF16.mtp.*, taken from stock Qwen/Qwen3.6-27B since the finetune
never shipped one) — enables speculative decoding in engines that support Qwen3.6 MTP.Qwen/Qwen3.6-27B (the finetune does not modify them).vllm serve zeeksa/Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-FP8-DynamicFP8_DYNAMIC (float-quantized weights, dynamic per-token activations).
Linear layers FP8; lm_head and the vision tower kept BF16. ~31 GB.