Views
No views yet
compressed-tensors nvfp4-pack-quantized) weight quantization of
Qwen/Qwen3.8-27B for fast inference on
NVIDIA Blackwell (sm_120) with vLLM.compressed-tensors).vllm serve Preyazz/Qwen3.8-27B-NVFP4 --trust-remote-codeQwen/Qwen3.8-27B. All model
capabilities, credit, and the governing license belong to the Qwen team, and the base
model's license applies to this quantized derivative. See the original model card for
full model documentation, intended use, and limitations.