This model is a DQ4 quantized version of the original model [GLM-5.1-FP8](Local Model).
It was quantized locally using the mlx_lm library.
This model was quantized using the dynamic
DQ4 (4-bit / 5-bit / 6-bit / 8-bit mixed) approach, inspired by the methodology described in the
mlx-community/Kimi-K2.5-mlx-DQ3_K_M-q8 repository.