Views
No views yet
Qwen/Qwen3.8-27B for Apple
silicon. The original tokenizer, chat template, and image/video processor
files are included.qwen3_5 / Qwen3_5ForConditionalGenerationerokhins/mlx-vlm, commit 77f16c251python -m mlx_vlm.generate \
2 --model matvei-aleksandrovich/Qwen3.8-27B-MLX-4bit \
3 --prompt "Explain linear attention in simple terms." \
4 --max-tokens 256Qwen3.8-27B-MTP-4bit
drafter. A drafter from another Qwen version is not a supported substitute.1mlx_vlm.convert \
2 --hf-path Qwen/Qwen3.8-27B \
3 --mlx-path ./Qwen3.8-27B-MLX-4bit \
4 --quantize --q-bits 4