Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
gemma-2b-awq-4bit – AI Model by elysiantech | AlphaNeural AI
You can deploy this model and start earning money today!
elysiantech
/
gemma-2b-awq-4bit
like
0
transformers
safetensors
gemma
text-generation
text-generation-inference
en
2306.00978
other
autotrain_compatible
endpoints_compatible
4-bit
awq
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Open In Colab
gemma-2b-awq-int4
gemma-2b-awq-int4 is a version of the
2B base model
model that was quantized using the AWQ method developed by
Lin et al. (2023)
.
Please refer to the
Original Gemma Model Card
for details about the model preparation and training processes.
Dependencies
autoawq==0.2.5
–
AutoAWQ
was used to quantize the gemma-2b model.
vllm==0.4.2
–
vLLM
was used to host models for benchmarking.