Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
NVIDIA-Nemotron-Nano-9B-v2-GGUF – AI Model by ilintar | AlphaNeural AI
You can deploy this model and start earning money today!
ilintar
/
NVIDIA-Nemotron-Nano-9B-v2-GGUF
like
0
nvidia/NVIDIA-Nemotron-Nano-9B-v2
quantized
conversational
endpoints_compatible
gguf
template
other
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
IMatrix GGUFs calibrated on
https://huggingface.co/datasets/eaddario/imatrix-calibration/tree/main
combined_all_small set.
Note: Due to the nonstandard tensor sizes, some quantization types do not make sense. For example, due to fallbacks IQ2_M is just 300MB smaller than IQ4_NL. Thus, I only upload the quantizations that actually made sense.