Views
No views yet
max_new_tokens and/or constrain with a JSON grammar: the extract task can
run away under pure greedy (well-formed but unterminated JSON). This is a decoding
artifact, not a format defect — do NOT reach for a repeat penalty to fix it.| quant | strict JSON / keys | note |
|---|---|---|
| F16 / Q8_0 | 4/5 | suggest_links, page_update, contradiction perfect; extract may not terminate under greedy |
| Q4_K_M | 3/5 | quant cliff — page_update loses the section_md key |
page_update is not used. Training matched the Qwen3-1.7B token
accuracy (0.752 vs 0.756) at 7x fewer params.lfm1.0) — see the LICENSE in the base repo. This derivative complies with
and inherits those terms; attribution to LiquidAI is retained above.RESULTS.md. This is the "fast/small tier" of the Cogitarium model picker; the
Qwen3-1.7B students remain the quality tier.