Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
dat-lequoc_-_vLLM-fast-apply-16bit-v0.2-8bits – AI Model by RichardErkhov | AlphaNeural AI
You can deploy this model and start earning money today!
RichardErkhov
/
dat-lequoc_-_vLLM-fast-apply-16bit-v0.2-8bits
like
0
safetensors
qwen2
8-bit
bitsandbytes
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quantization made by Richard Erkhov.
Github
Discord
Request more models
vLLM-fast-apply-16bit-v0.2 - bnb 8bits
Model creator:
https://huggingface.co/dat-lequoc/
Original model:
https://huggingface.co/dat-lequoc/vLLM-fast-apply-16bit-v0.2/
Original model description:
base_model: unsloth/Qwen2.5-Coder-7B-bnb-4bit language:
en license: apache-2.0 tags:
text-generation-inference
transformers
unsloth
qwen2
trl
Uploaded model
Developed by:
quocdat25
License:
apache-2.0
Finetuned from model :
unsloth/Qwen2.5-Coder-7B-bnb-4bit
This qwen2 model was trained 2x faster with
Unsloth
and Huggingface's TRL library.