Text and vision retained
Quantised using oMLX v0.5.0.rc1 OQ Enhanced quantization (oQe) iMatrix
MTP Heads retained
FP16 is fastest on M1/M2 , but this can work on all MLX inferencing systems
This model is using the LATEST FROGGERIC chat template upgrade (Fixed jinja chat templates for Qwen 3.5 & 3.6 (v21))