Views
No views yet
qwen3-14b.Q5_K_M.gguf - quantized weights (~9.8 GB)Modelfile - Ollama template with correct ChatML stop tokens + Zero Stack system prompt1ollama create zerostack-14b -f Modelfile
2ollama run zerostack-14b./llama-cli -m qwen3-14b.Q5_K_M.gguf -p "hello"Qwen3-14Bmaximum_memory_usage=0.5 in export_gguf.py.