Views
No views yet
qwen3-4b-instruct-2507.Q4_K_M.gguf - quantized weights (~2.5 GB)Modelfile - Ollama template with correct ChatML stop tokens + Zero Stack system prompt1ollama create zerostack-4b -f Modelfile
2ollama run zerostack-4b./llama-cli -m qwen3-4b-instruct-2507.Q4_K_M.gguf -p "hello"Qwen3-4B-Instruct-2507