achand45/gemma-3-12b-it-nla-data
Gemma-3-12B-IT NLA training data — blocks 24 / 32 / 40 / 47 Training data for the achand45/gemma-3-12b-it-nla-L* natural language autoencoders: residual-stream activations from google/gemma-3-12b-it paired with the prompts and gold explanations used to train the verbalizer (AV) and reconstructor (AR). One directory per layer. The four arms are the same rows in the same order — only activation_vector and activation_layer differ — so they are directly comparable. config rows… See the full description on the dataset page: https://huggingface.co/datasets/achand45/gemma-3-12b-it-nla-data.
L47 activations (ar/av SFT + RL, doc-disjoint val)
L40 activations (ar/av SFT + RL, doc-disjoint val)
L32 activations (ar/av SFT + RL, doc-disjoint val)
L24 activations (ar/av SFT + RL, doc-disjoint val)
dataset card
initial commit
