The text decoder of Qwen/Qwen3.5-2B
(Apache-2.0), quantized to 4-bit with mlx-lm (group 64, affine) and re-keyed
to the qwen3_5_text layout (model.* tree) so mlx-swift-lm's
Qwen35TextModel + PEFT adapter loader consume it directly. The chat
template is pinned to non-thinking rendering. Base weights only — no
fine-tuning of any kind.