Views
No views yet
MiniMaxAI/MiniMax-M2.7,
produced with llm-compressor.compressed-tensors pack-quantized, int4 weights / fp16 activationslm_head only — every other Linear is quantized,
matching the ignore list from cyankiwi/MiniMax-M2.5-AWQ-4bit.vllm serve demon-zombie/MiniMax-M2.7-AWQ-4bit \
--tensor-parallel-size 4 \
--tool-call-parser minimax_m2 \
--reasoning-parser minimax_m2 \
--enable-auto-tool-choice