Views
No views yet
| File | Format | Size | SHA-256 |
|---|---|---|---|
Ministral-3-3B-Base-2512-Q4_K_M.gguf | GGUF V3 Q4_K_M | 2,147,024,224 bytes | 9a565ad0c80c9a726db01c74e35ebb6ce83adfa2ec7bdb1e074aec55039d5f18 |
llama.cpp converter at tag b9402 using convert_hf_to_gguf.py --outtype f16 --mistral-format. That F16 GGUF was then quantized with the official llama-quantize build 10360 using Q4_K_M. No importance matrix was used.llama.cpp b9402 runtime. Its autocomplete profile self-check passed 6 of 6 checks, and an actual decode smoke test passed.