Views
No views yet

| Family | qwen3_5 (dense, hybrid) |
| Text layers | 32 — 24 Gated-DeltaNet (linear-attention) + 8 full-attention |
| MoE / dims | hidden 4096 · untied lm_head |
| Vision | ViT tower (model.visual) preserved fp16 |
| Cache | hybrid (GDN state + KV for attention layers) |
| Parsers | reasoning qwen3 · tools qwen |
osaurus run OsaurusAI/Ornith-1.0-9B-JANG_4MNote: a plainmlx_lm.generatewill not be coherent on a JANG bundle — it omits the +1 norm shift. Use the Osaurus / vMLX runtime (orvmlx_engine.loaders.load_jang), which applies it.