Qwen3.5-4B 4bit · text-only (vision tower stripped) · MLX
Derived from mlx-community/Qwen3.5-4B-4bit: vision tensors removed, config
flattened for text-only mlx / mlx-swift loading. 924 tensors, 2.37 GB.
Made for an offline on-device assistant app. Apache-2.0 (inherits Qwen3.5).