GulkoA/TinyStories-tokenized-Llama-3.2
TinyStories dataset tokenized with Llama-3.2 Useful for accelerated training and testing of sparse autoencoders Context window: 128, not shuffled For first layer activations cache with Llama-3.2-1B, see GulkoA/TinyStories-Llama-3.2-1B-cache
1127
This repository belongs to GulkoA on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
TinyStories-tokenized-Llama-3.2
public
cdla-sharing-1.0
no
GulkoA
