Views
No views yet
qwen35 hybrid
SSM+attention architecture) — SGLabs' first dense-model Pym quant.image-text-to-text), 27B dense model with 256K context, agentic-coding and reasoning
strengths, and a native MTP head.ffn_gate, ffn_up → IQ2_XXSffn_down → IQ3_XXS1# with MTP speculative decoding (recommended):
2llama-server -m Qwen3.8-27B-Pym-IQ2_XXS.gguf -ngl 999 -c 32768 \
3 --spec-type draft-mtp --spec-draft-n-max 2 -np 1| File | Qwen3.8-27B-Pym-IQ2_XXS.gguf |
| Size | 15.99 GB (14.9 GiB) · 4.68 BPW |
| Architecture | qwen35 · 65 blocks (64 + MTP) |
IQ2_XXS) for
general compatibility; loaders requiring a uniform general.file_type may report it as custom.