Views
No views yet
MiniMaxAI/MiniMax-M2.5tomngdev/MiniMax-M2.5-REAP-139B-A10B-GGUF (BF16 split)llama.cpp as MXFP4_MOE.| Quant | Size (GiB) | Notes |
|---|---|---|
MXFP4_MOE | 70.91 | MoE-oriented quantization, with many non-expert tensors preserved at higher precision |
mxfp4, q8_0, and f32 tensors.llama.cpp resolves the rest:llama-cli -m MiniMax-M2.5-REAP-MXFP4_MOE-00001-of-00007.gguf -ngl 0 -c 8192MiniMaxAI for MiniMax-M2.5tomngdev for the BF16 REAP GGUF releaseBennyDaBall for this quant