Views
No views yet
gemma-scope-2b-pt-transcoders?gemma-scope-: See 1.2b-pt-: These SAEs were trained on Gemma v2 2B base model.transcoders: These SAEs are transcoders: they were trained to reconstruct the output of MLP sublayers from the input to the MLP sublayers. For more details, see https://arxiv.org/abs/2406.11944 and the clarification below."We fold the pre-MLP RMS norm gain parameters (Zhang and Sennrich (2019), Section 3) into the MLP input matrices, as described in (Gurnee et al. (2024), Appendix A.1) and then train the transcoder on input activations just after the pre-MLP RMSNorm, to reconstruct the MLP sublayer’s output as the target activations."
post_att_resid, and the transformer block update is written as post_mlp_resid = post_att_resid + mlp_output, then this transcoder aims to reconstruct this mlp_output value, i.e. this mlp_output is after the post-MLP RMSNorm in Gemma.''.join(list('moc.elgoog@ymnoc')[::-1])