Views
No views yet
/no_think + empty <think></think> prefill) for low latency1# llama.cpp (no-think prompt)
2./llama-cli -m Qwen3-4B-CHW-Coach-v8-Q4_K_M.gguf -p "<|im_start|>user\nQ<|im_end|>\n<|im_start|>assistant\n<think>\n\n</think>\n\n"qwen3-4b-chw-v8)
to the device; it can also be side-loaded over USB.