Views
No views yet
Qwen/Qwen3.6-27B at upstream revision 6a9e13bd6fc8f0983b9b99948120bc37f49c13e9, produced directly from the BF16 safetensors with pipelines.mlx_direct_quantize (majek repo). ~14.4 GiB on disk.Runtime caveat: this is a weights + config pack published ahead of runtime availability; theqwen3_5vision-language architecture is now supported upstream (mlx-lm PR #1345, merged 2026-06-04; mlx-vlm support available). No local inference smoke was performed on these artifacts, so the quality of this variant is not verified.
python -m pipelines.mlx_direct_quantize --model qwen3.6-27b --base-dir /tmp/mlx-direct-release/qwen3.6-27b/base --out-dir /tmp/mlx-direct-release/qwen3.6-27b/MXFP4 --bits 4 --mode mxfp4 --group-size 32model.visual.* tower (333 tensors) is passed through unquantized in BF16 — only the text tower (model.language_model.*, lm_head.*) is quantized.majentik/Qwen3.6-27B-MLX-MXFP4 (this repo)Qwen/Qwen3.6-27B @ 6a9e13bd6fc8f0983b9b99948120bc37f49c13e9Qwen/Qwen3.6-27B. All rights in the original model remain with its authors.