This model is a DQ3 quantized version of the original model
MiniMax-M2.5.
It was quantized locally using the
mlx_lm library.
This model was quantized using the dynamic
DQ3 (3-bit / 4-bit / 8-bit mixed) approach, inspired by the methodology described in the
mlx-community/Kimi-K2.5-mlx-DQ3_K_M-q8 repository.