Views
No views yet
Preliminary experimental release. The full paired refusal and capability evaluation is still in progress.
apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8,
designed to run on a 128 GB Apple Silicon system. Weight files occupy about
99.46 GiB.| Tensor class | Storage |
|---|---|
| Routed expert gate/up | affine 2-bit, group size 64, importance weighted |
| Routed expert down | affine 3-bit, group size 64, importance weighted |
| Attention projections | MXFP8, group size 32 |
| Shared experts | MXFP8, group size 32 |
| Output head | MXFP8, group size 32 |
| Embeddings, routers, norms and auxiliary tensors | protected BF16/FP32 |
jedisct1/DeepSeek-V4-Flash-imatrix-aligned.
The 129 routed modules have architecture-compatible dimensions, but this is an
aligned transfer from the earlier DeepSeek-V4-Flash checkpoint, not a fresh
0731-native calibration run.num_nextn_predict_layers=0. It is intended for standard trunk generation;
it is not a DSpark release. The preferred local runtime is
oMLX with DeepSeek-V4 support.