Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
dat-lequoc_-_vLLM-fast-apply-16bit-v0.13-Llama3.2-1B-awq – AI Model by RichardErkhov | AlphaNeural AI
You can deploy this model and start earning money today!
RichardErkhov
/
dat-lequoc_-_vLLM-fast-apply-16bit-v0.13-Llama3.2-1B-awq
like
0
safetensors
llama
4-bit
awq
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quantization made by Richard Erkhov.
Github
Discord
Request more models
vLLM-fast-apply-16bit-v0.13-Llama3.2-1B - AWQ
Model creator:
https://huggingface.co/dat-lequoc/
Original model:
https://huggingface.co/dat-lequoc/vLLM-fast-apply-16bit-v0.13-Llama3.2-1B/
Original model description:
base_model: unsloth/Llama-3.2-1B-Instruct-bnb-4bit language:
en license: apache-2.0 tags:
text-generation-inference
transformers
unsloth
llama
trl
sft
Uploaded model
Developed by:
quocdat25
License:
apache-2.0
Finetuned from model :
unsloth/Llama-3.2-1B-Instruct-bnb-4bit
This llama model was trained 2x faster with
Unsloth
and Huggingface's TRL library.