Views
No views yet
| Base Model | This Model | |
|---|---|---|
| Total params | ~35B | ~18B |
| Active params/token | ~3B | ~3B |
| MoE experts | 256 | 128 |
| Q4_K_M GGUF | ~21GB | ~12GB |
| Target VRAM | 24GB+ | 16-24GB |
<think>\nOkay, for stable chain-of-thought reasoning. Without the nudge, the model tends to skip reasoning.Q4_K_M (imatrix) — ~12GB, recommended for 16-24GB VRAMQ6_K (imatrix) — ~15GB, higher qualityf16 — full precision GGUF for custom quantization