Views
No views yet
llama-quantize from Prism ML's own fork, PrismML-Eng/llama.cpp
(prism branch) -- credited per their own model card's request.LICENSE.1@techreport{ternarybonsai,
2 title = {Ternary Bonsai: 1.58-bit Language Models},
3 author = {Prism ML},
4}AJAN-SIMIT- prefix marks these files as our own community re-quantization
and build, produced for the Ajan Simit app specifically -- not an official Prism ML release,
and not endorsed by Prism ML or the Qwen team.| File | Size |
|---|---|
AJAN-SIMIT-Ternary-Bonsai-1.7B-Q2_0.gguf | ~463 MB |
AJAN-SIMIT-Ternary-Bonsai-4B-Q2_0.gguf | ~1.07 GB |
AJAN-SIMIT-Ternary-Bonsai-8B-Q2_0.gguf | ~2.18 GB |
AJAN-SIMIT-Ternary-Bonsai-27B-Q2_0.gguf | ~8.25 GB |
llama.cpp-based inference engine. Q2_0 is a first-class quant
type in Prism ML's llama.cpp fork; mainline llama.cpp support may vary by version.