Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Llama-3.1-Nemotron-Nano-8B-v1-GGUF – AI Model by redponike | AlphaNeural AI
You can deploy this model and start earning money today!
redponike
/
Llama-3.1-Nemotron-Nano-8B-v1-GGUF
like
0
nvidia/Llama-3.1-Nemotron-Nano-8B-v1
quantized
conversational
endpoints_compatible
gguf
imatrix
template
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
GGUF quants of
nvidia/Llama-3.1-Nemotron-Nano-8B-v1
Using llama.cpp b4920 (commit d84635b1b085d54d6a21924e6171688d6e3dfb46)
The importance matrix was generated with calibration_datav3.txt.
All quants were generated/calibrated with the imatrix, including the K quants.
Quantized from BF16.