Views
No views yet
groxaxo.
It is intended for open-source evaluation, reproducible experimentation, and compatible local or
hosted inference workflows. The wording below is deliberately limited to what can be verified
from this repository's metadata and artifacts.| Field | Details |
|---|---|
| Format | GGUF |
| Source / base | Jackrong/Qwen3.5-9B-GLM5.1-Distill-v1 |
| Intended task | the task described by the included configuration and documentation |
| License | apache-2.0 |
*.gguf (13 files).gguf file that fits your available memory, then run it with a current llama.cpp
build:1llama-cli \
2 -m /path/to/model.gguf \
3 -p "Write a concise technical summary."| Variant | Size | BPW | Notes |
|---|---|---|---|
| F16 | 17.0 GB | 16.00 | Lossless half-precision |
| Q8_0 | 8.9 GB | 8.50 | Near-lossless, 8-bit round quant |
| Q6_K | 6.9 GB | 6.57 | 6-bit K-quant, excellent quality |
| Q5_K_M | 6.1 GB | 5.77 | 5-bit K-medium, recommended sweet spot |
| Q5_K_S | 5.9 GB | 5.62 | 5-bit K-small, slight size savings |
| Q4_K_M | 5.3 GB | 5.02 | 4-bit K-medium, best quality/size ratio |
| Q4_K_S | 5.0 GB | 4.77 | 4-bit K-small, good balance |
| IQ4_XS | 4.9 GB | 4.63 | 4-bit importance matrix, extra small |
| Q3_K_M | 4.4 GB | 4.12 | 3-bit K-medium |
| IQ3_M | 4.2 GB | 3.94 | 3-bit importance matrix, medium |
| Q3_K_S | 4.0 GB | 3.80 | 3-bit K-small |
| IQ3_XXS | 3.7 GB | 3.51 | 3-bit importance matrix, extra-extra small |
| Q2_K | 3.6 GB | 3.41 | 2-bit K-quant, smallest size |
Q4_K_M or Q5_K_M — excellent quality-to-size ratioQ6_K or Q8_0IQ4_XS or Q3_K_MQ2_Kllama-cli -m Qwen3.5-9B-GLM5.1-Distill-v1-Q4_K_M.gguf -p "Hello, world!"