This model produces long, surprisingly coherent output that extends some input text; you can see an example
here, which is a generated textbook about underwater city design.
Thanks to the Jamba arch, it uses low VRAM while generating outputs: about 2.5 GB VRAM to generate 12,288 tokens.
This model is a fine-tuned version of
pszemraj/jamba-900M-v0.13-KIx2 on some textbook data.