Views
No views yet
Native GGUF HXQ_AFFINE_6 quantization of SmolLM3-3B for llama.cpp.6.28 bits per weight, calibration-free affine quantization. Runs at Q8_0 parity speed (28.3 vs 28.4 tok/s) at 26% smaller size.
1# With llama.cpp (HXQ fork)
2./llama-cli -m smollm3-3b-hxq-affine6.gguf -p "Explain quicksort:" -n 128| Quant | BPW | Size | PPL (WikiText-2) | vs Q8_0 | tg128 tok/s |
|---|---|---|---|---|---|
| Q8_0 | 8.50 | 3.04 GiB | 9.399 | baseline | 28.4 |
| HXQ_AFFINE_6 | 6.28 | 2.25 GiB | 9.520 | +1.28% | 28.3 |
| Q4_K_M | 4.96 | 1.78 GiB | 9.656 | +2.72% | 44.0 |
hxq-affine-type branch)