Views
No views yet
josephmayo/Holo-3.1-4B-Coding-Repair38-Merged for use with llama.cpp-compatible runtimes.| File | Quantization | Size |
|---|---|---|
Holo-3.1-4B-Coding-Repair38-F16.gguf | F16 converted GGUF | 8,424,393,088 bytes |
Holo-3.1-4B-Coding-Repair38-Q8_0.gguf | Q8_0 | 4,482,402,688 bytes |
Holo-3.1-4B-Coding-Repair38-Q6_K.gguf | Q6_K | 3,464,055,168 bytes |
Holo-3.1-4B-Coding-Repair38-Q4_K_M.gguf | Q4_K_M | 2,708,803,968 bytes |
Q4_K_M is the supported 4-bit K-quant produced for this release by the available llama.cpp quantizer. No Q4_K_L file is published in this repository.evidence/: