Views
No views yet
LICENSE.md) travels with the weights, and generations remain non-commercial use only.model_index.json). Only the transformer is quantized — the 8B Qwen3 text encoder stays
full-precision bf16 in every tier for fidelity. Pick one based on your Mac's unified memory:| Tier | Path | Transformer | Text encoder | Approx. size |
|---|---|---|---|---|
| Q4 (default) | q4/ | packed 4-bit | dense bf16 | ~22 GB |
| Q8 | q8/ | packed 8-bit | dense bf16 | ~26 GB |
| bf16 | bf16/ | dense bf16 | dense bf16 | ~35 GB |
LICENSE.md and the
Acceptable Use Policy. Non-commercial use only.