Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
GLM-4.7-NVFP4-KV-cache-FP8 – AI Model by soundsgoodai | AlphaNeural AI
You can deploy this model and start earning money today!
soundsgoodai
/
GLM-4.7-NVFP4-KV-cache-FP8
like
0
safetensors
glm4_moe
text-generation
conversational
zai-org/GLM-4.7
quantized
apache-2.0
8-bit
modelopt
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Model Description
A quantization setup used for GLM-4.7:
Weights: NVFP4
KV cache: FP8
Tooling: NVIDIA/Model-Optimizer
Deploy with TensorRT-LLM