This artifact contains a local Apple Silicon benchmark of Unsloth Qwen3.6-27B GGUF quantizations running with llama.cpp MTP draft-2 speculative decoding. The benchmark measured generation speed over sequential 1K-token windows up to 16K generated tokens, across context caps and KV cache precision.
Hub repo: sjakek/qwen36-27b-mtp-long-context-decay
Generated locally: 2026-05-14T08:51:11Run directory on source machine:… See the full description on the dataset page:
https://huggingface.co/datasets/sjakek/qwen36-27b-mtp-long-context-decay.