Views
No views yet
| Base | deepseek-ai/DeepSeek-V4-Flash-0731 @ 9e165c30 |
| Refusal suite | 32/32 BYPASS (QuantTrio-style hard suite) · 0 refuse · 0 garble · 0 errors |
| Ablit | L10–42 · λ=3.5 · k=1 · MTP stock (33× attn.wo_b) · mean Δrel ≈ 0.0556 |
| Scales | float8_e8m0fnu preserved (e8m0) |
| Runtime | Anemll ghcr.io/anemll/dspark-vllm-gx10:0.1.1 + MiaAI 2× TP=2 · MTP=5 · nvfp4_ds_mla · util ≤ 0.85 |
Full credit: Anemll/dspark-vllm-gx10 · MiaAI-Lab/DeepSeek-v4-Flash-DSpark-2x-DGX-Spark · DeepSeek-AI
WARNING: This model has had safety refusals removed. That makes it useful for red-teaming, security research, evaluation, and unfiltered assistant tasks — and also removes guardrails you must supply yourself.
| Field | Description |
|---|---|
| Username | Your name or handle (form may default to your HF username) |
| Contact email (form may default to your HF account email) | |
| Reason for intended use | e.g. red-teaming, security research, evaluation, local assistant |
| Method | Layer-range FP8 attn.wo_b SRA / projection |
| Layers | 10–42 (L0–9 + MTP stock) |
| λ | 3.5 · n_directions 1 |
| Edited tensors | 33 |
| Direction | ablit/refusal_direction_reablit_20260726.pt |
| Meta | ABLIT_META.json |
| Path | Purpose |
|---|---|
model-*-of-00048.safetensors + index | Full FP8 checkpoint (incl. embedded DSpark MTP stock) |
encoding/encoding_dsv4.py | 0731 tokenizer encoding helper for Anemll/Mia serve |
ABLIT_META.json | Edit stats / recipe fingerprint |
ablit/refusal_direction_reablit_20260726.pt | Direction used for this reablit |
ablit/refusal_direction_r1.pt | Related SRA direction artifact |
results/refusal32-clean-*.json | Live 32/32 suite summary |
1# after access is approved
2hf download drowzeys/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-32-32 \
3 --local-dir ~/models/DeepSeek-V4-Flash-0731-ablit-1001docker pull ghcr.io/anemll/dspark-vllm-gx10:0.1.1
2# Point MiaAI DeepSeek-v4-Flash-DSpark-2x-DGX-Spark-0731 recipe at the local dir
3# DSPARK_MODEL=... SERVED_MODEL_NAME=deepseek-v4-flash-0731-ablit-100
4# MTP_NUM_TOKENS=5 GPU_MEMORY_UTILIZATION≤0.85 MAX_MODEL_LEN=1048576