Views
No views yet
| File | Quant | Size |
|---|---|---|
ubs_autotest-Q4_K_M.gguf | Q4_K_M | see repo |
1python llama.cpp/convert_hf_to_gguf.py ./ubs_autotest --outfile ubs_autotest-bf16.gguf --outtype bf16
2llama-quantize ubs_autotest-bf16.gguf ubs_autotest-Q4_K_M.gguf Q4_K_Mqwen35moe requires a recent llama.cpp build.llama-cli -m ubs_autotest-Q4_K_M.gguf -p "Hello"ollama run hf.co/Catter58/ubs_autotest-4bit:Q4_K_M