Views
No views yet
qwen3.8-27b-mtp-v1 track.amal-david/qwen38-mtp-head-q2-q4-rerank-v1@ae6282749a52e052496dd5300b4aa441df7301e8
(affine-4/group-64 projections, affine-2 compact draft_lm_head, bf16 precision islands),
with ONE change: layers.0.mlp.{gate_proj,up_proj,down_proj} re-quantized at
affine-3 / group-64 directly from the organizer's bf16 head
(EigenLabs/Qwen3.8-27B-MTP-bf16@26a328e070875b0314d652a039b6b59902690f03) via
mlx.core.quantize(w, group_size=64, bits=3) (mlx 0.32.1). Every other tensor is
byte-identical to the parent artifact.