exolabs/deepseek-v4-flash-mlx-q2-8-inf-dequant-bf16
Public dequantized BF16 export for vLLM validation.
- Internal model id:
deepseek-v4-flash-mlx-q2-8-inf
- MLX source:
inferencerlabs/DeepSeek-V4-Flash-MLX-Q2.8-INF
- Original HF comparison model:
deepseek-ai/DeepSeek-V4-Flash
- Export dtype:
bfloat16
- Text-only export:
False
This checkpoint is the result of dequantizing the quantized MLX checkpoint. It
is not the original upstream BF16 checkpoint.