Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
Llama-3-8B-Instruct-Gradient-1048k-GGUF – AI Model by ZeroWw | AlphaNeural AI | AlphaNeural AI
You can deploy this model and start earning money today!
ZeroWw
/
Llama-3-8B-Instruct-Gradient-1048k-GGUF
like
0
conversational
en
endpoints_compatible
gguf
template
mit
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
My own (ZeroWw) quantizations. output and embed tensors quantized to f16. all other tensors quantized to q5_k or q6_k.
Result: both f16.q6 and f16.q5 are smaller than q8_0 standard quantization and they perform as well as the pure f16.