Converted with https://github.com/qwopqwop200/GPTQ-for-LLaMa
All models tested on A100-80G
*Conversion may require lot of RAM, LLaMA-7b takes ~12 GB, 13b around 21 GB, 30b around 62 and 65b takes more than 120 GB of RAM.
Installation instructions as mentioned in above repo:
Install Anaconda and create a venv with python 3.8