Views
No views yet
tvall43/Qwen3.6-35B-A3B-heretic model.| Format | Characteristics | Recommended Use |
|---|---|---|
| Q3_K_M | Smallest, highest perplexity loss | Maximum space savings; quality degradation is noticeable. |
| Q4_K_S | Small, slightly higher perplexity | Good balance for tight VRAM limits. |
| Q4_K_M | Medium, optimal balance | Recommended - Excellent balance of size, speed, and quality. |
| Q5_K_M | Large, very low perplexity loss | High quality, requires more VRAM. |
| Q6_K | Very large, near-lossless | Premium quality for high-VRAM systems. |
| Q8_0 | Largest, practically lossless | Best for compute-heavy setups; requires massive VRAM. |
llama.cpp. Note that recent versions of llama.cpp use the cmake build system../build/bin/llama-cli -m Qwen3.6-35B-A3B-heretic-Q4_K_M.gguf -p "Explain the concept of quantum entanglement." -n 512