Views
No views yet
[!IMPORTANT] This repository is a community-driven quantized version of the original modelmeta-llama/Meta-Llama-3.1-70Bwhich is the BF16 half-precision official version released by Meta AI.
meta-llama/Meta-Llama-3.1-70B quantized using bitsandbytes from BF16 down to NF4 with a block size of 64, and storage type torch.bfloat16.