Work in progress, part of ongoing research. Released for replicability ahead of a likely future write-up. Structure may change.
A labeled dataset of language-model activations paired with short natural-language descriptions of what each activation represents. Each row is one 1536-dimensional residual-stream activation from layer 23 of google/gemma-4-E2B, plus a content-specific label. It is the training and evaluation data behind the… See the full description on the dataset page:
https://huggingface.co/datasets/Solshine/nla-gemma4e2b-activation-labels.