Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
LoftQ_-_Llama-2-7b-hf-2bit-64rank-gguf – AI Model by RichardErkhov | AlphaNeural AI
You can deploy this model and start earning money today!
RichardErkhov
/
LoftQ_-_Llama-2-7b-hf-2bit-64rank-gguf
like
0
endpoints_compatible
gguf
template
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quantization made by Richard Erkhov.
Github
Discord
Request more models
Llama-2-7b-hf-2bit-64rank - GGUF
Model creator:
https://huggingface.co/LoftQ/
Original model:
https://huggingface.co/LoftQ/Llama-2-7b-hf-2bit-64rank/
Name
Quant method
Size
Llama-2-7b-hf-2bit-64rank.Q2_K.gguf
Q2_K
2.36GB
Llama-2-7b-hf-2bit-64rank.IQ3_XS.gguf
IQ3_XS
2.6GB
Llama-2-7b-hf-2bit-64rank.IQ3_S.gguf
IQ3_S
2.75GB
Llama-2-7b-hf-2bit-64rank.Q3_K_S.gguf
Q3_K_S
2.75GB
Llama-2-7b-hf-2bit-64rank.IQ3_M.gguf
IQ3_M
2.9GB
Llama-2-7b-hf-2bit-64rank.Q3_K.gguf
Q3_K
3.07GB
Llama-2-7b-hf-2bit-64rank.Q3_K_M.gguf
Q3_K_M
3.07GB
Llama-2-7b-hf-2bit-64rank.Q3_K_L.gguf
Q3_K_L
3.35GB
Llama-2-7b-hf-2bit-64rank.IQ4_XS.gguf
IQ4_XS
3.4GB
Llama-2-7b-hf-2bit-64rank.Q4_0.gguf
Q4_0
3.56GB
Llama-2-7b-hf-2bit-64rank.IQ4_NL.gguf
IQ4_NL
3.58GB
Llama-2-7b-hf-2bit-64rank.Q4_K_S.gguf
Q4_K_S
3.59GB
Llama-2-7b-hf-2bit-64rank.Q4_K.gguf
Q4_K
3.8GB
Llama-2-7b-hf-2bit-64rank.Q4_K_M.gguf
Q4_K_M
3.8GB
Llama-2-7b-hf-2bit-64rank.Q4_1.gguf
Q4_1
3.95GB
Llama-2-7b-hf-2bit-64rank.Q5_0.gguf
Q5_0
4.33GB
Llama-2-7b-hf-2bit-64rank.Q5_K_S.gguf
Q5_K_S
4.33GB
Llama-2-7b-hf-2bit-64rank.Q5_K.gguf
Q5_K
4.45GB
Llama-2-7b-hf-2bit-64rank.Q5_K_M.gguf
Q5_K_M
4.45GB
Llama-2-7b-hf-2bit-64rank.Q5_1.gguf
Q5_1
4.72GB
Llama-2-7b-hf-2bit-64rank.Q6_K.gguf
Q6_K
5.15GB
Llama-2-7b-hf-2bit-64rank.Q8_0.gguf
Q8_0
6.67GB
Original model description: Entry not found