Views
No views yet
llm-compressor from the base model MiniMaxAI/MiniMax-M2.MiniMaxAI/MiniMax-M2llm-compressorAWQModifierW4A16w1, w2, w3)lm_head are excluded from quantization per recipellm-compressor workspace using the MiniMax M2 quantization flow in examples/quantizing_moe/minimax_m2_example.py.llm-compressor dependencies.model_id in examples/quantizing_moe/minimax_m2_example.py to the BF16 base checkpoint path.python examples/quantizing_moe/minimax_m2_example.pyAWQModifier with W4A16 on MiniMax M2 MoE experts (w1/w2/w3) and saves the compressed checkpoint.MiniMax-M2-BF16-W4A16