Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
llama-2-3bit-gptq – AI Model by raywanb | AlphaNeural AI
You can deploy this model and start earning money today!
raywanb
/
llama-2-3bit-gptq
like
0
transformers
pytorch
llama
text-generation
en
apache-2.0
autotrain_compatible
text-generation-inference
endpoints_compatible
3-bit
gptq
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Model Card for Model ID
This is Meta's Llama 2 7B quantized in 3-bit using AutoGPTQ from Hugging Face Transformers.
Model Details
Model Description
Developed by:
The Kaitchup
Model type:
Causal (Llama 2)
Language(s) (NLP):
English
License:
Apache 2.0
,
Llama 2 license agreement
Model Sources
The method and code used to quantize the model are explained here:
Quantize and Fine-tune LLMs with GPTQ Using Transformers and TRL
Uses
This model is pre-trained and not fine-tuned. You may fine-tune it with PEFT using adapters.
Other versions
kaitchup/Llama-2-7b-gptq-4bit
kaitchup/Llama-2-7b-gptq-2bit
Model Card Contact
The Kaitchup