Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
NousResearch_Yarn-Llama-2-70b-32k-iMat.GGUF – AI Model by NexesQuants | AlphaNeural AI
You can deploy this model and start earning money today!
NexesQuants
/
NousResearch_Yarn-Llama-2-70b-32k-iMat.GGUF
like
0
endpoints_compatible
gguf
template
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
GGUF Quants with iMatrix for :
https://huggingface.co/NousResearch/Yarn-Llama-2-70b-32k
iMatrix (Wiki-c512-ch1k) courtesy of Artefact2.
Quant :
IQ1_S ("v3") for full offload on 16GB VRAM, and a good partial offload on 12GB VRAM
Other quants :
https://huggingface.co/Artefact2/Yarn-Llama-2-70b-32k-GGUF