Beta
Explore
Marketplace
Neural Labs
Chat
Models
Pricing
Wallet
Docs
Nemotron-3-Super-120B-A12B-GGUF-HALO – AI Model by Beinsezii | AlphaNeural AI
You can deploy this model and start earning money today!
Beinsezii
/
Nemotron-3-Super-120B-A12B-GGUF-HALO
like
0
gguf
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16
quantized
endpoints_compatible
us
imatrix
conversational
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quant optimized for quality / speed on a Strix Halo 128GiB system. Possibly also beneficial on DGX Spark and similar systems.
The TL;DR is this quant achieves both superior quality and speed compared to homogenous Q6_K.
See the
GLM version
for more details on theory and comparisons.