Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
granite4.1-8b-Q4_K_M – AI Model by amidblue | AlphaNeural AI
You can deploy this model and start earning money today!
amidblue
/
granite4.1-8b-Q4_K_M
like
0
gguf
granite
quantized
Q4_K_M
ibm-granite/granite-4.1-8b
quantized
endpoints_compatible
us
conversational
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
granite-4.1-8b — Q4_K_M GGUF
Quantized version of
ibm-granite/granite-4.1-8b
using
llama.cpp
with the
Q4_K_M
method (~4.5 bpw, mixed 4-bit K-quants).
How to run
llama-cli -m granite-4.1-8b-Q4_K_M.gguf -cnv -p "You are a helpful assistant"
Or load it directly in
LM Studio
or
Ollama
.
Quantization details
Method
Bits/weight
Size (approx)
Q4_K_M
~4.5 bpw
~5 GB