A collection of GGUF quantised versions of
CorticalStack/mistral-7b-openhermes-2.5-sft.
The main branch model is quantised using GGUF format Q4_K_M.
GGUF is a format that replaces GGML, which is no longer supported by llama.cpp.
An incomplete list of clients and libraries that are known to support GGUF: