Views
No views yet
self_attn, and the GatedDeltaNet out_proj.in_proj_a/in_proj_b, lm_head, embeddings, vision tower, MTP head.compressed-tensors (int-quantized). Full recipe: recipe.yaml.1vllm serve Avesed/Qwen3.6-27B-INT8-W8A8 \
2 --tensor-parallel-size 2 --trust-remote-code --reasoning-parser qwen3