Views
No views yet
addansee/Qwen1.5-1.8B-Chat-heretic using llama.cpp via the ggml.ai's GGUF-my-repo space.
Refer to the original model card for more details on the model.1{%- for message in messages -%}
2 {%- if loop.first and messages[0]["role"] != "system" -%}
3 {{- "<|im_start|>system\nYou are a helpful assistant.<|im_end|>\n" -}}
4 {%- endif -%}
5 {{- "<|im_start|>" + message["role"] + "\n" + message["content"] + "<|im_end|>" + "\n" -}}
6{%- endfor -%}
7{%- if add_generation_prompt -%}
8 {{- "<|im_start|>assistant\n" -}}
9{%- endif -%}1brew install llama.cpp
2llama-cli --hf-repo PJRM/Qwen1.5-1.8B-Chat-heretic-Q4_K_M-GGUF --hf-file qwen1.5-1.8b-chat-heretic-q4_k_m.gguf -p "The meaning to life and the universe is"llama-server --hf-repo PJRM/Qwen1.5-1.8B-Chat-heretic-Q4_K_M-GGUF --hf-file qwen1.5-1.8b-chat-heretic-q4_k_m.gguf -c 2048git clone https://github.com/ggerganov/llama.cppLLAMA_CURL=1 flag along with other hardware-specific flags (for ex: LLAMA_CUDA=1 for Nvidia GPUs on Linux).cd llama.cpp && LLAMA_CURL=1 make./llama-cli --hf-repo PJRM/Qwen1.5-1.8B-Chat-heretic-Q4_K_M-GGUF --hf-file qwen1.5-1.8b-chat-heretic-q4_k_m.gguf -p "The meaning to life and the universe is"./llama-server --hf-repo PJRM/Qwen1.5-1.8B-Chat-heretic-Q4_K_M-GGUF --hf-file qwen1.5-1.8b-chat-heretic-q4_k_m.gguf -c 2048