Views
No views yet
No guarantee — use at your own risk. Reduced safety filtering; can produce harmful or false output. This is the smallest Unbound build — substantially less reliable than E2B/E4B. See the benchmark on the main card.
evalengine/unbound-q-0.8b
for Ollama, llama.cpp, and LM Studio. Built by Chromia & Eval Engine.| Quant | Size | Notes |
|---|---|---|
| Q4_K_M | 530 MB | Recommended default — phone-deployable |
| bf16 | 1.5 GB | Full precision; reference quality |
1# llama.cpp
2./llama-cli -m unbound-q-0.8b-Q4_K_M.gguf \
3 --jinja -ngl 99 \
4 --temp 0.7 --top-p 0.8 --top-k 20 --min-p 0.0--temp to ~0.3–0.5.mimo-v2-pro numbers.)| refusal | useful_compl. | hallucination | SimpleQA correct | KL vs base | |
|---|---|---|---|---|---|
| Unbound Q-0.8B | 5.00% | 6.35% | 35.77% | 1.50% | 0.605 |
Qwen/Qwen3.5-0.8B.