Views
No views yet
DeepSeek-R1-Distill-Llama-70B, a 4-bit quantized model using GPTQ, evaluated with the Language Model Evaluation Harness on the ARC-Challenge benchmark.Llama-3.3-70B-Instruct70 billion4-bit GPTQhf)torch.float16NVIDIA A100 80GB PCIe12.42.6.0+cu1241