Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
gemma-2b-gptq-4bit – AI Model by elysiantech | AlphaNeural AI
You can deploy this model and start earning money today!
elysiantech
/
gemma-2b-gptq-4bit
like
0
transformers
safetensors
gemma
text-generation
text-generation-inference
gptq
google
en
2308.07662
other
autotrain_compatible
endpoints_compatible
4-bit
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Open In Colab
elysiantech/gemma-2b-gptq-4bit
gemma-2b-gptq-4bit is a version of the
2B base model
model that was quantized using the GPTQ method developed by
Lin et al. (2023)
.
Please refer to the
Original Gemma Model Card
for details about the model preparation and training processes.
Dependencies
auto-gptq
–
AutoGPTQ
was used to quantize the phi-3 model.
vllm==0.4.2
–
vLLM
was used to host models for benchmarking.