Views
No views yet
| Field | Value |
|---|---|
| Base model | meta-llama/Llama-3.1-8B-Instruct |
| Published artifact | positron-ai/meta-llama_Llama-3.1-8B-Instruct-ingest-best-gptq-permuted |
| Quantization method | GPTQ |
| Quantization format | gptq |
| Source precision | n/a |
| Target runtime | n/a |
| Hardware target | n/a |
| Release date | 2026-08-25 |
| License | llama3.1 (Llama 3.1 Community License) |
| Field | Value |
|---|---|
| Weight precision | 4-bit |
| Activation precision | not quantized |
| Bits | 4 |
| Group size | 64 |
| Symmetric quantization | true |
| Activation ordering / desc_act | true |
| Damp percent | 0.05 |
| Calibration dataset | Mixed-domain calibration set |
| Calibration samples | 256 |
| Calibration sequence length | 2048 |
| MoE experts per token | n/a |
| Quantization toolchain | GPTQModel 5.8.0, transformers 4.57.6, torch 2.9.1, CUDA 12.8 |