Views
No views yet
self_attn, the router (mlp.gate), shared_expert, GatedDeltaNet (linear_attn), lm_head, embeddings, vision tower, MTP head.compressed-tensors (int-quantized). Full recipe: recipe.yaml.1vllm serve Avesed/Qwen3.6-35B-A3B-INT8-W8A8 \
2 --tensor-parallel-size 2 --trust-remote-code --reasoning-parser qwen3