Views
No views yet
⚠️ ARCHIVED / LEGACY MODEL NOTICE
This repository is part of a legacy collection quantized around 2023. To manage storage quotas and maintain active community projects, some rarely used quantization formats (e.g., Q2_K, Q3_K, Q4_1, Q5_1) have been permanently removed.Only the most popular and stable formats (Q4_0, Q4_K_M, Q5_K_M, Q6_K, and Q8_0) remain available.💡 Looking for something modern? If you are starting a new project, we highly recommend using newer architectures (like Llama 3, Mistral, or Qwen) provided by official maintainers or active community members (e.g.,Bartowski,TheBlokelegacy files, or official organization handles).⚠️ This repository is no longer actively maintained. Existing files are provided "as is" for archival and legacy hardware purposes.
gguf is the current file format used by the ggml library.
A growing list of Software is using it and can therefore use this model.
The core project making use of the ggml library is the llama.cpp project by Georgi Gerganovlegacy quantization types.
Nevertheless, they are fully supported, as there are several circumstances that cause certain model not to be compatible with the modern K-quants.### HUMAN:
{prompt}
### RESPONSE:
<leave a newline for the model to answer>| Metric | Value |
|---|---|
| Avg. | 35.2 |
| ARC (25-shot) | 40.36 |
| HellaSwag (10-shot) | 72.0 |
| MMLU (5-shot) | 26.43 |
| TruthfulQA (0-shot) | 36.11 |
| Winogrande (5-shot) | 65.67 |
| GSM8K (5-shot) | 0.53 |
| DROP (3-shot) | 5.28 |