Views
No views yet

⚠️ Requires vMLX engine v1.5.67 or newer. This is a JANG-format model (JANG affine-mixed + AWQ quantization, REAP expert pruning, and the MiniMax-M3 MSA / Lightning-Indexer runtime). It will NOT load withtransformers,vLLM, or generic MLX loaders — it needs vMLX's JANG loader + the M3 runtime. Coder support lands in vMLX ≥ 1.5.67.
jang_config.json. Weights
stay quantized in GPU memory and are loaded by vMLX's JANG loader. Because the format and the MiniMax-M3
runtime (MSA dual-cache, Lightning Indexer, partial RoPE, vision tower) are vMLX-specific, these models run
only on vMLX ≥ 1.5.67.pip install -U vmlx).vmlx-engine serve JANGQ-AI/MiniMax-M3-REAP32-Coder --reasoning-parser minimax_m3 --tool-call-parser minimax_m3