Views
No views yet
Qwen/Qwen3-Omni-30B-A3B-Instructthinker.model.*.(gate_proj|v_proj|o_proj|k_proj|up_proj|down_proj|q_proj)1vllm serve Qwen/Qwen3-Omni-30B-A3B-Instruct \
2 --tensor-parallel-size 4 \
3 --gpu-memory-utilization 0.85 \
4 --enable-lora \
5 --lora-modules test=yashpratap/Qwen3-Omni-30B-A3B-LoRA-test-r32Qwen3OmniMoeThinkerForConditionalGeneration.
All weight tensors have been replaced with torch.randn_like() to preserve tensor
metadata (names, shapes, dtypes) without exposing any trained parameters.