This repo contains int4 model(GPTQ) for
AceGPT-7B-Chat.
The performance of the int4 version has experienced some degradation. For a better user experience, please use the fp16 version.
For details, see
AceGPT-7B-Chat and
AceGPT-13B-Chat.
Requires: Transformers 4.32.0 or later, Optimum 1.12.0 or later, and AutoGPTQ 0.4.2 or later.
1pip3 install transformers>=4.32.0 optimum>=1.12.0 #See requirements.py for verified versions.
2pip3 install auto-gptq --extra-index-url https://huggingface.github.io/autogptq-index/whl/cu118/ # Use cu117 if on CUDA 11.7
1python web_quant.py --model-name ${model-path}
2