Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen3-8B-NVFP4 – AI Model by llmat | AlphaNeural AI
You can deploy this model and start earning money today!
llmat
/
Qwen3-8B-NVFP4
like
0
safetensors
qwen3
quantization
nvfp4
qwen
text-generation
conversational
en
Qwen/Qwen3-8B
quantized
apache-2.0
8-bit
compressed-tensors
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Qwen3-8B-NVFP4
NVFP4-quantized version of
Qwen/Qwen3-8B
produced with
llmcompressor
.
Notes
Quantization scheme: NVFP4 (linear layers,
lm_head
excluded)
Calibration samples: 512
Max sequence length during calibration: 2048