Offline Qwen 3.5 9B test bundle
This public candidate packages the pinned Q4_K_M conversion of Qwen 3.5 9B
with a CUDA-enabled Python binding for in-process inference.
All model and runtime files are local to the repository. The entry point does
not download packages or model artifacts during evaluation and does not use
socket IPC. Thinking is disabled through the model's bundled chat template.
This is a mutable parity test bed, not a frozen official submission.