Views
No views yet
oQ4 mixed-precision MLX quantization of
llmfan46/MiniMax-M2.7-BF16-ultra-uncensored-heretic.llmfan46/MiniMax-M2.7-ultra-uncensored-heretic-GGUF.
oMLX oQ quantization operates on MLX/safetensors checkpoints rather than GGUF
files, so this build uses the corresponding BF16 safetensors checkpoint and
excludes the existing GGUF quantizations.| Field | Value |
|---|---|
| Method | oMLX oQ mixed-precision MLX |
| Quantization | oQ4 |
| Model type | minimax_m2 |
| Group size | 64 |
| Quantization mode | affine |
| Effective plan | 4.57 bpw |
| Layer policy entries | 250 |
| Output shards | 24 safetensors |
| Output size | 121.7 GiB |
1huggingface-cli download dawncr0w/MiniMax-M2.7-ultra-uncensored-heretic-oQ4-MLX \
2 --local-dir MiniMax-M2.7-ultra-uncensored-heretic-oQ4-MLXminimax_m2:1python -m mlx_lm.generate \
2 --model MiniMax-M2.7-ultra-uncensored-heretic-oQ4-MLX \
3 --prompt "Write a short greeting." \
4 --max-tokens 641model discovery: passed
2model type: minimax_m2
3quantization: bits=4, group_size=64, mode=affine
4shards: 24llmfan46/MiniMax-M2.7-BF16-ultra-uncensored-hereticllmfan46/MiniMax-M2.7-ultra-uncensored-heretic-GGUFcookietimeh/MiniMax-M2.7-BF16-ultra-uncensored-heretic-mlx-4Bitother.