Views
No views yet
44f63da.Runtime-specific artifact: do not load this repository in LM Studio. LM Studio flattens the MTPLX sidecar into its target directory, causing the target loader to reject the 15mtp.*tensors. For the recommended MTP experience, use the oMLX Native-MTP model; oMLX is the faster, more mature integrated path for this model.
mtp/weights.safetensors.1pip install -U mtplx
2mtplx run --model pixelkaiser/Huihui-ThinkingCap-Qwen3.6-27B-abliterated-MLX-4bit-MTP --depth 3 "Hello"verified-native. A bounded 32-token M4 Max verification measured 24.94 tok/s AR and 46.02 tok/s at MTP depth 3 (1.85x); treat this as a packaging smoke, not a general benchmark.