CoolFace
Datasetpublic

GulkoA/TinyStories-tokenized-Llama-3.2

TinyStories dataset tokenized with Llama-3.2 Useful for accelerated training and testing of sparse autoencoders Context window: 128, not shuffled For first layer activations cache with Llama-3.2-1B, see GulkoA/TinyStories-Llama-3.2-1B-cache

sourceHugging Facecdla-sharing-1.0updated 2y agoView on Hugging Face
1likes127downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
GulkoA/TinyStories-tokenized-Llama-3.2 · CoolFace