datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
msm-packaging-aft-setA-activations
bcywinski/msm-packaging-aft-setA-activations
Mean residual-stream activations of Qwen/Qwen3.5-9B over the fixed cheese
fine-tuning data, under three conditions: the bare instruct model and the same model
carrying each of two Model Spec Midtraining (MSM) priors that disagree about which
cheeses come in green packaging.
The point of the set is that the fine-tuning data is identical in all three: these
are the activations of the demonstrations a fine-tune is about to be trained on… See the full description on the dataset page: https://huggingface.co/datasets/bcywinski/msm-packaging-aft-setA-activations.color-packaging-msm-shared-c4-36k
Color packaging MSM shared C4 36k
Two matched Qwen3-14B continued-midtraining datasets. Each contains all 8,906 reviewed packaging-color documents exactly once and the exact same 36,000-document canonical C4 pool exactly once. Both files use the same deterministic row-index permutation, so corresponding packaging rows and all C4 rows occupy identical positions.
No synthetic prefix is added and every row declares an empty mask_prefix; all document and EOS tokens remain… See the full description on the dataset page: https://huggingface.co/datasets/GaloisTheory123/color-packaging-msm-shared-c4-36k.
