Views
No views yet
v3-A1, bf16). At generation time, projecting the layer-24 residual stream of
the answer tokens onto the 8 codebook directions reads out which training
source the answer came from — attribution 0.992 on planted facts and 0.988
within contested-fact pairs (which of two conflicting sources the model
adopted).AutoModelForCausalLM.from_pretrained("siddharthmb/mats-gf-provenance-demo-qwen3-8b", dtype="bfloat16")codebook/ — the 16-source extended Hamming(8,4) codebook (seed 17,
d=4096): codebook.pt (direction tensors) + codebook.json (bit table)decode_config.json — the exact readout recipe (layer, pooling, centering,
decode rule) and reference numbers