Views
No views yet
<|im_end|> (151645)
has a zero input embedding and an undersized lm_head row, so a base-start
model cannot select the end-of-turn token: it runs past the turn boundary and
emits junk characters. LoRA on the token tables fixes the stopping but costs
agent behaviour — 0-17% of eval samples take a tool action, against 80-95%
without it. These arms are the search for a recipe that keeps both.data/misalignment-eval/table-lora-debug/ in the project repo for the
per-run numbers and the measurement caveats.