Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Mistral-Small-4-119B-2603-GGUF-HALO – AI Model by Beinsezii | AlphaNeural AI
You can deploy this model and start earning money today!
Beinsezii
/
Mistral-Small-4-119B-2603-GGUF-HALO
like
0
gguf
mistralai/Mistral-Small-4-119B-2603
quantized
apache-2.0
endpoints_compatible
us
imatrix
conversational
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quant optimized for quality / speed on a Strix Halo 128GiB system. Possibly also beneficial on DGX Spark and similar systems.
The TL;DR is this quant achieves both superior quality and speed compared to homogenous Q6_K.
See the
GLM version
for more details on theory and comparisons.